Senior Software Engineer - AI Evaluation / Coding Agents
Senior Software Engineer - AI Evaluation / Coding Agents is an open role at Turing AI Advancement Work, available remotely.
Applications open on Turing AI Advancement Work’s own site. How we make money.
About this role
Freelance · Remote · North America, LATAM, or India About Turing Turing is one of the world’s leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent. About the Role We’re looking for experienced, hands-on software engineers to help evaluate and improve AI coding models . Rather than primarily building production applications, you’ll work with coding agents across real-world repositories and assess the quality of their work. You’ll review generated code and agent behavior, determine whether solutions are technically correct, identify failure modes, and create the evaluation signals and feedback used to improve model performance. Think of the coding agent as another engineer whose work you’re reviewing: Can it understand the task? Did it choose the right approach? Is the resulting code correct, robust, and maintainable? Can you explain precisely where it succeeded or failed? What You’ll Do Evaluate AI-generated code and solutions across real-world software repositories Review agent behavior, tool usage, and code changes for correctness and quality Identify technical errors, weak approaches, and recurring model failure modes Compare model outputs and explain why one solution is better than another Create and refine rubrics and evaluation criteria for coding tasks Produce high-quality evaluation and preference data used to improve coding models Build and maintain pipelines and infrastructure supporting data generation, collection, and evaluation workflows Synthesize findings from data work into clear write-ups, updates, and recommendations for the team Collaborate closely with researchers and engineers to translate qualitative judgment into scalable processes Share clear, actionable findings with AI researchers and engineers What We’re Looking For 5+ years of hands-on software engineering experience Strong proficiency in Python, TypeScript/JavaScript, Go, or another major production language Experience working in substantial real-world codebases Strong code-review skills and technical judgment Ability to clearly explain why an implementation is correct, incorrect, or could be improved Strong written communication Experience using modern LLMs or AI coding tools Experience with LLM evaluation, coding agents, RLHF, preference data, rubric design, or post-training is a plus, but not required. Engagement Details Compensation: Market rate; please provide a specific hourly rate expectation Availability: 40 hours/week preferred, with at least 6 hours of Pacific Time overlap Type: Independent contractor Duration: Approximately 3 months Start: As soon as possible Location: North America, LATAM, or India Evaluation Process AI interview (~25 minutes) Practical code/AI evaluation exercise (~30 minutes) Hiring manager interview (~20 minutes) The practical exercise focuses on your ability to review and evaluate AI-generated code , not competitive programming or algorithm puzzles.
Pay
Not disclosed
Turing AI Advancement Work doesn’t publish a rate for this listing - it isn’t missing from our data.
About Turing AI Advancement Work
Turing AI Advancement Work's connection to Turing: Direct project marketplace. Coding, data science, model evaluation, voice and domain review
Ownership confidence: Confirmed. See the full network profile →
Similar open roles
Frequently asked questions
How do I apply for the Senior Software Engineer - AI Evaluation / Coding Agents role at Turing AI Advancement Work?
Applications go through Turing AI Advancement Work's own portal - HumanSourcer links to the listing and never collects applications or handles hiring. Turing AI Advancement Work's access model is "Apply" (Apply to listed projects). This listing was first seen here on September 10, 2026 and was still live at the last check.
Is the Senior Software Engineer - AI Evaluation / Coding Agents role at Turing AI Advancement Work remote?
Yes - Turing AI Advancement Work lists this role as remote.
What does the Senior Software Engineer - AI Evaluation / Coding Agents role at Turing AI Advancement Work pay?
Turing AI Advancement Work doesn't state a rate on this listing. HumanSourcer never fills that gap with an estimate, so the absence is the source's, not a hole in this page. Turing AI Advancement Work's profile shows what its rate-carrying listings do advertise.
What kind of work is Senior Software Engineer - AI Evaluation / Coding Agents?
Coding / software engineering. Coding, data science, model evaluation, voice and domain review
Is Turing AI Advancement Work a legitimate AI-training platform?
Direct project marketplace - and that relationship is rated "Confirmed" here because it can be checked against work.turing.com rather than taken on Turing AI Advancement Work's word. Confidence describes how well the ownership is evidenced. It is not a rating of how the platform treats the people who work for it, which no public source covers reliably.
Get new AI-training roles by email
One weekly digest of newly listed roles across every tracked network - pay where it's disclosed, no spam, unsubscribe anytime.
Free. See how we make money.