thinking machines labMillbrae, CA
AI Safety Researcher: Safety, Evaluation & Red-Teaming
AI Safety Researcher: Safety, Evaluation & Red-Teaming
AI Safety Researcher: Safety, Evaluation & Red-Teaming
thinking machines labMillbrae, CA
yesterday
Computer and Information Research ScientistsHealth and Safety Engineers, Except Mining Safety Engineers and InspectorsData Scientists
Other Scientific and Technical Consulting ServicesResearch and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)Software Publishers
Apply for this role →Thinking Machines Lab Inc. in San Francisco is seeking a safety researcher to bridge research and hands-on engineering, focusing on making models safe and trustworthy.
You will explore how training shapes refusals and how to evaluate boundaries, designing experiments to inform model training and evaluation strategies.
Join a team that values rigorous analysis, data-driven evaluation, and responsible AI practices, with opportunities to influence real-world deployments.
Also on the board Same function, level within a rung
Level
Junior
Location
Millbrae, CA
Occupation
Computer and Information Research Scientists
Industry
Other Scientific and Technical Consulting Services
Posted
yesterday