You’ll run agentic coding sessions on real engineering tasks, analyze what the models produce, and red-team them until they fail in interesting ways. Every failure you document teaches the next generation of coding agents.
What you’ll actually do
- Vibe-code with frontier models on real tasks: fix a bug, extend an API, ship a feature, and see how far the model gets on its own.
- Analyze the AI’s work against production standards: find the bugs, explain the failure, and rate the model’s reasoning.
- Red-team the models to expose unsafe code, faked test passes, and confidently wrong solutions before real users hit them.
Roles this fits
Common backgrounds: Software Engineer, Backend Engineer, Full-Stack Developer, DevOps Engineer.
What we look for
- Professional or serious open-source experience shipping production code.
- Comfort in at least one major stack; most tasks use Python, C++, Rust, Go, or JavaScript.
- Clear written English: your explanations are the training signal.
- No degree required. We care about what you can do, not where you learned it.
Steps Involved
- Apply
- Qualify
- Work & get paid
Compensation
Up to $40 – $150+/hr depending on task difficulty and specialization. Many contributors add $10k–$100k+ a year; some make it their full-time income.
About DataAnnotation
DataAnnotation is where 100k+ experts train the world’s leading AI models. $150M+ paid to contributors to date, and the average contributor stays 5+ years. Flexible, remote, and always project-available.