Skip to content
HumanSourcer
Turing AI Advancement Work

Senior Python Engineer – AI Agents Evaluation

Coding / software engineeringTuring AI Advancement WorkCodingRemote

Senior Python Engineer – AI Agents Evaluation is an open role at Turing AI Advancement Work, available remotely.

Apply at Turing AI Advancement Work →

Applications open on Turing AI Advancement Work’s own site. How we make money.

About this role

About Turing Turing is one of the world's leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent. We are staffing a frontier AI data initiative building the infrastructure and training data used to develop and evaluate AI agents. You will be assigned to one of two tracks: connectors or tasks. About the Role We are looking for experienced Python Engineers to work with leading AI research labs on evaluating and improving next-generation AI coding agents. You will review agentic coding trajectories, assess technical correctness, debug and validate outputs, and use coding agents such as Claude Code, Codex, or Cursor as part of your workflow. What You'll Do Review agentic coding trajectories, including prompts, tool calls, intermediate actions, code changes, execution results, and final outputs. Evaluate whether AI agents correctly understand and execute software engineering tasks. Analyze agent behavior to identify technical errors, inefficient approaches, and opportunities for improvement. Read, debug, test, and validate Python code produced or modified by AI agents. Work with LLM-powered agents, workflows, and development environments to evaluate model capabilities. Use coding agents such as Claude Code, Codex, Cursor, or similar tools to accelerate analysis, debugging, and validation. Requirements 5+ years of professional software engineering experience, with strong hands-on expertise in Python. At least 6 months to 1 year of practical AI/LLM engineering experience building agents, agent loops, LLM-powered applications, data pipelines, or similar systems. Hands-on experience using AI coding agents such as Claude Code, Codex, Cursor, or equivalent tools as part of regular software development workflows. Ability to understand and evaluate agentic workflows, including tool usage, intermediate execution steps, code modifications, and resulting outputs. Perks of Working With Turing Work with leading AI research labs and contribute to the development of cutting-edge AI systems. Fully remote opportunity. Collaborate with a global network of talented engineers and AI professionals. Flexible engagement with exposure to frontier AI development and evaluation. Offer Details Engagement: Full-time contractual opportunity Location: Remote Availability: 40 hours per week with required overlap with the global team Contract Duration: 1 month Evaluation process Around 25 mins of AI interview

Pay

Not disclosed

Turing AI Advancement Work doesn’t publish a rate for this listing - it isn’t missing from our data.

Share this listing on RedditShare this listing on XShare this listing on LinkedIn

About Turing AI Advancement Work

Turing AI Advancement Work's connection to Turing: Direct project marketplace. Coding, data science, model evaluation, voice and domain review

Ownership confidence: Confirmed. See the full network profile →

Similar open roles

Frequently asked questions

How do I apply for the Senior Python Engineer – AI Agents Evaluation role at Turing AI Advancement Work?

Applications go through Turing AI Advancement Work's own portal - HumanSourcer links to the listing and never collects applications or handles hiring. Turing AI Advancement Work's access model is "Apply" (Apply to listed projects). This listing was first seen here on August 13, 2026 and was still live at the last check.

Is the Senior Python Engineer – AI Agents Evaluation role at Turing AI Advancement Work remote?

Yes - Turing AI Advancement Work lists this role as remote.

What does the Senior Python Engineer – AI Agents Evaluation role at Turing AI Advancement Work pay?

Turing AI Advancement Work doesn't state a rate on this listing. HumanSourcer never fills that gap with an estimate, so the absence is the source's, not a hole in this page. Turing AI Advancement Work's profile shows what its rate-carrying listings do advertise.

What kind of work is Senior Python Engineer – AI Agents Evaluation?

Coding / software engineering. Coding, data science, model evaluation, voice and domain review

Is Turing AI Advancement Work a legitimate AI-training platform?

Direct project marketplace - and that relationship is rated "Confirmed" here because it can be checked against work.turing.com rather than taken on Turing AI Advancement Work's word. Confidence describes how well the ownership is evidenced. It is not a rating of how the platform treats the people who work for it, which no public source covers reliably.

Get new AI-training roles by email

One weekly digest of newly listed roles across every tracked network - pay where it's disclosed, no spam, unsubscribe anytime.

Free. See how we make money.

Apply at Turing AI Advancement Work →