Thinking Machines Lab Inc. in San Francisco is seeking a safety researcher to bridge research and hands-on engineering, focusing on making models safe and trustworthy. You will explore how training shapes refusals and how to evaluate boundaries, designing experiments to inform model training and evaluation strategies. Join a team that values rigorous analysis, data-driven evaluation, and responsible AI practices, with opportunities to influence real-world deployments.

Also on the board Same function, level within a rung

Level

Junior

Location

Millbrae, CA

Occupation

Computer and Information Research Scientists

Industry

Other Scientific and Technical Consulting Services

Posted

yesterday

Apply for this role →