About the role

The newest models can vibe-code a whole feature in minutes. Someone has to check whether that code actually works, and that someone is you. You'll run agentic coding sessions on real engineering tasks, analyze what the models produce, and red-team them until they fail in interesting ways. Every failure you document teaches the next generation of coding agents. What you’ll actually do Vibe-code with frontier models on real tasks: fix a bug, extend an API, ship a feature, and see how far the model gets on its own. Analyze the AI's work against production standards: find the bugs, explain the failure, and rate the model's reasoning. Red-team the models to expose unsafe code, faked test passes, and confidently wrong solutions before real users hit them.
What we look for: Professional or serious open-source experience shipping production code. Clear written English: your explanations are the training signal. No degree required. We care about what you can do, not where you learned it. Compensation Up to $40 – $150+/hr depending on task difficulty and specialization. Many contributors add $10k–$100k+ a year; some make it their full-time income. About Data Annotation Data Annotation is where 100k+ experts train the world’s leading AI models. $150M+ paid to contributors to date, and the average contributor stays 5+ years. Flexible, remote, and always project-available.

Matching similar jobs

JOB OVERVIEW

Salary

$40 – $150+/hr

Experience level

Mid

Location

Ruidoso, NM

Occupation

Astronomers

Industry

All Other Professional, Scientific, and Technical Services

Posted

today

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE