The newest models can vibe-code a whole feature in minutes. Someone has to check whether that code actually works, and that someone is you. You'll run agentic coding sessions on real engineering tasks, analyze what the models produce, and red-team them until they fail in interesting ways. Every failure you document teaches the next generation of coding agents. What you’ll actually do Vibe-code with frontier modelson real tasks: fix a bug, extend an API, ship a feature, and see how far the model gets on its own. Analyze the AI's workagainst production standards: find the bugs, explain the failure, and rate the model's reasoning. Red-team the modelsto expose unsafe code, faked test passes, and confidently wrong solutions before real users hit them.
What we look for: Professional or serious open-source experience shipping production code. Clear written English: your explanations are the training signal. No degree required. We care about what you can do, not where you learned it. Compensation Up to $40 – $150+/hr depending on task difficulty and specialization. Many contributors add $10k–$100k+ a year; some make it their full-time income. About Data Annotation Data Annotation is where 100k+ experts train the world’s leading AI models. $150M+ paid to contributors to date, and the average contributor stays 5+ years. Flexible, remote, and always project-available.

Also on the board Same function, level within a rung

Level

Mid

Salary

$10,000 per month

Location

Salem, OH

Occupation

Bioinformatics Scientists

Industry

All Other Professional, Scientific, and Technical Services

Posted

yesterday

Apply for this role →