Mercor is building a red team to probe AI models with adversarial inputs and surface vulnerabilities. This role focuses on reviewing outputs touching sensitive topics and creating high-quality data to strengthen AI safety. You will conduct jailbreaking, prompt injection, and bias-exploitation tests across multi-turn conversations, documenting reproducible attack cases and delivering artifacts for customers.

Also on the board Same function, level within a rung

Level

Mid

Location

New York, NY

Occupation

Penetration Testers

Industry

Custom Computer Programming Services

Posted

5 days ago

Apply for this role →