Skip to content
HumanSourcer
Sepal Expert Hub

Physics PhD Reasoning Benchmark

AI model evaluationSepal Expert HubScience / STEMReasoning / AgentsGlobal

Physics PhD Reasoning Benchmark is an open role at Sepal Expert Hub, with pay listed as $65–$100/hr.

Apply at Sepal Expert Hub →

Applications open on Sepal Expert Hub’s own site. How we make money.

About this role

We're developing novel methods to evaluate how well AI systems can reason through advanced physics workflows. This project sits at the cutting edge of AI evaluation—where physics domain knowledge meets frontier model assessment. We're seeking Physics PhD holders (or ABDs) with strong conceptual depth and computational modeling experience to help us design and review high-quality physics problem sets that test AI’s capabilities in solving real-world physics challenges. What You'll Do - Design rigorous, graduate-level physics tasks and rubrics to assess AI reasoning across topics like classical mechanics, quantum systems, thermodynamics, and electrodynamics. - Contribute example solutions and grading rubrics that assess reasoning steps, not just final answers. - Collaborate with our internal team and other domain experts to probe the strengths and failure modes of advanced AI models. - Provide expert feedback on task quality, correctness, ambiguity, and conceptual coverage. What We're Looking For - PhD or ABD in Physics, Applied Physics, or a closely related discipline. - Strong ability to distill complex physics problems and guide others through the reasoning. - Demonstrated experience in computational modeling or numerical analysis (e.g., simulations, data pipelines, Monte Carlo methods, FEA). - Familiarity with at least one of: Python, MATLAB, Mathematica, Julia, C++, COMSOL, ROOT, Geant4, or equivalent scientific tooling. - Experience with publishing, teaching, or curriculum design is a plus. Ideal candidates will be comfortable with both problem design and conceptual review, especially for graduate-level tasks. Compensation & Logistics - Hourly Rate: $65–$100/hour depending on experience and task complexity. - Type: Contract, flexible hours (part-time and full-time options). - Location: Remote (US preferred); Bay Area availability is a bonus

Pay

$65–$100/hr

Share this listing on RedditShare this listing on XShare this listing on LinkedIn

About Sepal Expert Hub

Sepal Expert Hub's connection to Sepal AI: Direct expert network. Specialist AI training and evaluation

Ownership confidence: Confirmed. See the full network profile →

Similar open roles

Frequently asked questions

How do I apply for the Physics PhD Reasoning Benchmark role at Sepal Expert Hub?

Applications go through Sepal Expert Hub's own portal - HumanSourcer links to the listing and never collects applications or handles hiring. Sepal Expert Hub's access model is "Apply" (Apply to listed roles). This listing was first seen here on July 19, 2026 and was still live at the last check.

Is the Physics PhD Reasoning Benchmark role at Sepal Expert Hub remote?

This role is listed as open globally.

What does the Physics PhD Reasoning Benchmark role at Sepal Expert Hub pay?

Sepal Expert Hub lists this role's pay as $65–$100/hr, quoted from the listing rather than estimated.

What kind of work is Physics PhD Reasoning Benchmark?

AI model evaluation. Specialist AI training and evaluation

Is Sepal Expert Hub a legitimate AI-training platform?

Direct expert network - and that relationship is rated "Confirmed" here because it can be checked against expert-hub.sepalai.com rather than taken on Sepal Expert Hub's word. Confidence describes how well the ownership is evidenced. It is not a rating of how the platform treats the people who work for it, which no public source covers reliably.

Get new AI-training roles by email

One weekly digest of newly listed roles across every tracked network - pay where it's disclosed, no spam, unsubscribe anytime.

Free. See how we make money.

Apply at Sepal Expert Hub →