AI Developer Trace Task Auditor
Mercor (client confidential) · Remote — United States · Remote
- Pay
- $70–90/hr
- Commitment
- hourly
- Hours / week
- ~40
- Source
- mercor
About this role
Evaluate the quality and correctness of AI-assisted software-development traces used to train and evaluate a frontier AI lab's models. You'll assess end-to-end coding sessions produced with AI-assisted developer tools — judging correctness, workflow soundness, and reasoning — and provide clear, rubric-based written feedback. Basic Qualifications • 3+ years professional software development • Hands-on experience with AI-assisted coding tools and agentic / spec-driven workflows (Cursor, GitHub Copilot, Claude Code, or similar) • Strong code-reading and debugging skills across full-stack or backend systems • Ability to evaluate multi-step coding trajectories for correctness and best practice Preferred Qualifications • Experience with Kiro or Amazon CodeCatalyst • Prior work evaluating or grading AI-generated code • Contributions to developer tooling
Eligible applicant countries
This role accepts applicants from:
- USA
Skills & domains
- ai-training
- rlhf
- sme
- annotation
- Software Engineering
