Hillclimb is a research-focused startup aiming to improve agent capabilities toward recursive self-improvement. Our team designs environments and evals that push frontier models to ideate, experiment, and produce meaningful research progress. We value expertise in LLMs, RL, RLHF/RLAIF, and automated evaluation systems to ensure trustworthy rewards. This role emphasizes building scalable pipelines and validating environments before training, with a collaborative frontier/data lab setting.

Also on the board Same function, level within a rung

Level

Mid

Location

Millbrae, CA

Occupation

Computer and Information Research Scientists

Industry

Other Scientific and Technical Consulting Services

Posted

yesterday

Apply for this role →