Mercor is assembling a panel of chemistry and chemical safety experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones. You will write challenging prompts across benign, dual-use, and adversarial levels, evaluate the responses against a defined policy standard, and craft reference answers with the technical reasoning behind them.

Also on the board Same function, level within a rung

Level

Mid

Location

Millbrae, CA

Occupation

Occupational Health and Safety Specialists

Industry

Other Scientific and Technical Consulting Services

Posted

yesterday

Apply for this role →