
Applied Physics Benchmark Specialist | $61-$77/hr Remote
Overview
We are looking for seasoned physicists to craft and review advanced academic assessment content for an AI research project. You’ll develop and validate high quality multiple-choice questions across key physics areas, assess solution quality, and help establish benchmark materials that push AI capabilities forward. The role is remote and hourly, with flexible, asynchronous collaboration. You’ll work in two modes: authoring original questions and verifying pre-written items for accuracy and rigor.
What You'll Do8
- 1Create original physics questions that test deep understanding and avoid surface level recall
- 2Review and refine pre-written questions to improve clarity, precision, and solvability
- 3Assign difficulty levels such as Medium, Hard, or Expert and justify the rating
- 4Provide a correct answer plus 9 plausible distractors to challenge expert solvers
- 5Draft step-by-step solution outlines with clear intermediate steps
- 6Suggest 1–5 scholarly references per item from credible sources
- 7Flag issues in questions during verification and document edits made
- 8Support the development of gold-standard benchmarks used to advance AI capabilities
Requirements5
- 1PhD or doctoral candidate in Physics, Applied Physics, Astrophysics, or closely related fields
- 2Master’s degree considered for exceptional depth in a subdomain
- 3Strong grasp of graduate level physics concepts and mathematical formalism
- 4Experience in rigorous problem design or physics Olympiad style writing is a plus
- 5Excellent written English with the ability to convey complex ideas clearly
Who Should Apply
Ideal candidates are self-driven researchers with a deep physics background who enjoy crafting precise, challenging assessment items and can explain complex ideas succinctly. You should be comfortable working asynchronously, meeting quality benchmarks, and contributing to AI benchmarking efforts.
Salary Insight
Compensation is hourly in the range of $61.00 to $77.00 per hour.
Location
Required Skills
Application Tip
Prepare a concise sample question and accompanying solution outline that demonstrates your ability to define clear problem statements, create credible distractors, and justify the difficulty rating.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedApplied Engineering Benchmark Specialist | $61-$77/hr Remote
The role focuses on creating and validating rigorous academic assessment content for an AI research effort. You will author and review multiple-choice questions across core engineering topics, judge solution quality, and help build gold-standard benchmarks to push AI capabilities forward. This is a remote, hourly engagement, with tasks that blend technical evaluation and high-quality writing using a clear, structured approach.

Mercor
VerifiedApplied Computer Science Benchmark Specialist | $66-$84/hr Remote
A remote, hourly role focused on building and reviewing high quality AI research benchmark content for computer science. You’ll craft and validate rigorous multiple choice questions across core CS topics, assess answer quality, and contribute to gold-standard datasets that push AI capabilities forward. Expect a split between creating new questions and verifying existing ones, with opportunities to shape benchmarks used by researchers worldwide.

Mercor
VerifiedApplied Mathematics Benchmark Specialist | $61-$77/hr Remote
A remote role focused on creating and validating rigorous math assessment content for an AI research initiative. You’ll write and review advanced multiple choice questions across key mathematics domains, assess solution quality, and help set benchmarks that push AI capabilities forward. The work blends mathematical depth with clear, instructional problem design and documentation. This position stands out for its blend of academic rigor and practical benchmarking in AI research.

Mercor
VerifiedApplied Chemistry Benchmark Specialist | $61-$77/hr Remote
We’re looking for seasoned chemists to craft and review high quality academic assessment content for an AI research initiative. In this role you’ll develop and validate rigorous chemistry multiple choice questions, evaluate answer quality, and help set gold-standard benchmarks that push AI understanding forward. Work is remote and project-based, with flexible hours. You’ll work in areas spanning materials, polymer and electronic chemistry, industrial processes, energy storage, environmental chemistry, pharmaceutical and agrochemical chemistry, as well as consumer and food chemistry.

Mercor
VerifiedApplied Biology Benchmark Specialist | $60-$75/hr Remote
One sentence on its own. We’re looking for seasoned biologists to craft and review high quality academic assessment content for an AI research project. You’ll develop and verify rigorous biology multiple-choice items, assess solution quality, and help set gold-standard benchmarks that push AI capabilities forward. The role spans authoring original questions and validating existing ones across several biology domains using a remote, asynchronous setup and competitive hourly pay.

Mercor
VerifiedApplied Philosophy Benchmark Specialist | $50-$63/hr Remote
A remote role focused on creating and vetting high quality philosophy assessment content for an AI research initiative. You will craft and critique multiple choice questions across core philosophy domains, while helping set gold standard benchmarks for advancing AI capabilities. The work blends deep philosophical expertise with rigorous evaluation, and offers flexible, asynchronous engagement.