Skip to content
HumanSourcer
Turing AI Advancement Work

Engineering Expert

Data annotation / labelingTuring AI Advancement WorkScience / STEMRemote

Engineering Expert is an open role at Turing AI Advancement Work, available remotely.

Apply at Turing AI Advancement Work →

Applications open on Turing AI Advancement Work’s own site. How we make money.

About this role

About Turing: Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in coding, reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L. Role Overview We are seeking experienced AI Evaluation Engineers (Engineering Simulation & Design) to author and validate "model-breaking," simulation-based engineering design problems to train and evaluate state-of-the-art AI agents. Operating across major engineering disciplines—including Electrical, Mechanical, Control Systems, Aerospace, Systems, and Robotics—you will create complex, multi-constraint tasks where AI agents must interpret requirements, navigate trade-offs, configure open-source simulation tools, diagnose failures, and iterate toward valid solutions. You will analyze agent execution logs, expose systemic reasoning gaps, and build automated, objective graders to elevate frontier model performance. Job Requirements: Education & Expertise: Master’s degree or PhD in Electrical, Mechanical, Aerospace, with 10+ years of hands-on engineering design experience. Simulation Tooling: Proficiency with at least one domain-relevant open-source simulation package (e.g., ngspice, PySpice, OpenFOAM, FEniCSx, CalculiX, python-control, CadQuery, build123d, OpenModelica, Cantera, Gmsh) combined with strong Python scripting skills. AI Evaluation & Failure Diagnostics: Hands-on experience with modern LLMs/coding agents and evaluation concepts (pass@k, failure-mode analysis, nondeterministic behavior), with the ability to audit trajectory logs and isolate core reasoning/tool-use failures. Domain Rigor & Precision: Uncompromising attention to physical plausibility, unit consistency, boundary conditions, convergence criteria, and technical documentation. Availability & Commitment: Talent must have weekend on-call availability (part-time engagement is acceptable). Technical Infrastructure: Personal desktop/laptop equipped with a stable, high-speed internet connection in a remote setup. Job Responsibilities: Model-Breaking Problem Design: Author original, self-contained engineering design tasks with competing constraints, explicit optimization targets, validated reference solutions, and objective autograders. Environment & Simulation Integration: Build, run, and validate problem environments using open-source simulation tools and custom Python test benches. Trajectory Analysis & Failure Mode Taxonomy: Evaluate coding agent outputs and execution logs across repeated trials to identify systemic failure modes (e.g., misinterpreting simulator feedback, premature design convergence, physically impossible geometries). Difficulty Calibration & Benchmark Refinement: Iteratively refine problem difficulty based on empirical model performance data without introducing ambiguity or missing information. Cross-Functional Collaboration: Partner with AI researchers, pod leads, and domain experts to integrate high-rigor benchmarks into the model evaluation pipeline. Domains: Electrical Engineering Mechanical Engineering Aerospace Engineering Education & Experience Bachelor's degree or equivalent practical experience in any field. Experience in AI evaluation, data annotation, content review, quality assurance, or a related analytical role is preferred but not required. Offer Details: Commitments Required: 40 hours per week with 4 hours of overlap with PST. Engagement type: Contractor Engagement Length: upto 24 weeks Evaluation Process - Shortlisted candidates will be sent a Job Interest Form . Finalized talents will go through delivery review & proceeded further accordingly.

Pay

Not disclosed

Turing AI Advancement Work doesn’t publish a rate for this listing - it isn’t missing from our data.

Share this listing on RedditShare this listing on XShare this listing on LinkedIn

About Turing AI Advancement Work

Turing AI Advancement Work's connection to Turing: Direct project marketplace. Coding, data science, model evaluation, voice and domain review

Ownership confidence: Confirmed. See the full network profile →

Similar open roles

Frequently asked questions

How do I apply for the Engineering Expert role at Turing AI Advancement Work?

Applications go through Turing AI Advancement Work's own portal - HumanSourcer links to the listing and never collects applications or handles hiring. Turing AI Advancement Work's access model is "Apply" (Apply to listed projects). This listing was first seen here on September 10, 2026 and was still live at the last check.

Is the Engineering Expert role at Turing AI Advancement Work remote?

Yes - Turing AI Advancement Work lists this role as remote.

What does the Engineering Expert role at Turing AI Advancement Work pay?

Turing AI Advancement Work doesn't state a rate on this listing. HumanSourcer never fills that gap with an estimate, so the absence is the source's, not a hole in this page. Turing AI Advancement Work's profile shows what its rate-carrying listings do advertise.

What kind of work is Engineering Expert?

Data annotation / labeling. Coding, data science, model evaluation, voice and domain review

Is Turing AI Advancement Work a legitimate AI-training platform?

Direct project marketplace - and that relationship is rated "Confirmed" here because it can be checked against work.turing.com rather than taken on Turing AI Advancement Work's word. Confidence describes how well the ownership is evidenced. It is not a rating of how the platform treats the people who work for it, which no public source covers reliably.

Get new AI-training roles by email

One weekly digest of newly listed roles across every tracked network - pay where it's disclosed, no spam, unsubscribe anytime.

Free. See how we make money.

Apply at Turing AI Advancement Work →