
Applied Mathematics Benchmark Specialist | $61-$77/hr Remote
Overview
A remote role focused on creating and validating rigorous math assessment content for an AI research initiative. You’ll write and review advanced multiple choice questions across key mathematics domains, assess solution quality, and help set benchmarks that push AI capabilities forward. The work blends mathematical depth with clear, instructional problem design and documentation. This position stands out for its blend of academic rigor and practical benchmarking in AI research.
What You'll Do7
- 1Create original, challenging multi-step math questions that probe deep understanding and conceptual mastery
- 2Review existing questions for clarity, completeness, and precision, making edits as needed
- 3Assign difficulty levels such as Medium, Hard, or Expert to each item and document reasoning
- 4Develop a correct answer plus a set of 9 plausible distractors that are subtly misleading for experts
- 5Draft transparent, step-by-step solutions in markdown that illuminate the reasoning process
- 6Include 1–5 scholarly references per question from reputable academic sources
- 7For verification tasks, identify and justify issues related to clarity, completeness, precision, or solvability
Requirements5
- 1PhD or doctoral candidate in Mathematics, Applied Mathematics, Statistics, or a closely related field
- 2Exceptional depth in a specific subdomain can qualify a master’s degree
- 3Strong ability to reason through advanced mathematical concepts and formal proof writing
- 4Experience in rigorous problem design or competition-style writing is advantageous
- 5Excellent written English with the capacity to convey complex ideas concisely
Who Should Apply
Ideal candidates are deeply mathematical, enjoy crafting precise problems, and thrive in a research-driven setting. You should be comfortable articulating solutions clearly, can assess problem quality critically, and have a track record of contributing to high-level academic or competition-style material. Remote, asynchronous work fits well if you value independent collaboration and rigorous standards.
Salary Insight
Hourly compensation ranges from $61.00 to $77.00 per hour.
Location
Required Skills
Application Tip
Prepare a concise portfolio of 2–3 sample questions with solutions and a brief justification for their difficulty level to showcase your problem-design approach.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedApplied Computer Science Benchmark Specialist | $66-$84/hr Remote
A remote, hourly role focused on building and reviewing high quality AI research benchmark content for computer science. You’ll craft and validate rigorous multiple choice questions across core CS topics, assess answer quality, and contribute to gold-standard datasets that push AI capabilities forward. Expect a split between creating new questions and verifying existing ones, with opportunities to shape benchmarks used by researchers worldwide.

Mercor
VerifiedApplied Engineering Benchmark Specialist | $61-$77/hr Remote
The role focuses on creating and validating rigorous academic assessment content for an AI research effort. You will author and review multiple-choice questions across core engineering topics, judge solution quality, and help build gold-standard benchmarks to push AI capabilities forward. This is a remote, hourly engagement, with tasks that blend technical evaluation and high-quality writing using a clear, structured approach.

Micro1
VerifiedMathematics Expert | $80-$90/hr Remote
The Mathematics Expert will work remotely as a contractor to support a high-impact AI training project. You bring deep mathematical knowledge to produce rigorous problems, proofs, and explanations that guide model learning. This role centers on research-driven content creation, clear communication, and collaboration with a global team.

Micro1
VerifiedMathematics Expert | $20-$40/hr Remote
Mathematics Expert role available for remote contractors. You’ll apply deep mathematical knowledge to support training of AI systems, shaping how models reason and perform on real-world data. This position values clear explanations, rigorous reasoning, and precise communication over casual background experience. You’ll work with a distributed team to deliver high-quality insights and datasets that drive model benchmarking and improvement.

Mercor
VerifiedApplied Physics Benchmark Specialist | $61-$77/hr Remote
We are looking for seasoned physicists to craft and review advanced academic assessment content for an AI research project. You’ll develop and validate high quality multiple-choice questions across key physics areas, assess solution quality, and help establish benchmark materials that push AI capabilities forward. The role is remote and hourly, with flexible, asynchronous collaboration. You’ll work in two modes: authoring original questions and verifying pre-written items for accuracy and rigor.

Mercor
VerifiedApplied Health & Medicine Benchmark Specialist | $94-$119/hr Remote
We’re looking for seasoned health science experts to create and review high quality assessment content for an AI research effort. You’ll craft and validate multiple choice questions across core medical domains, assess solution quality, and help set gold standard benchmarks to push AI capabilities forward. Work is remote and on an hourly basis, with flexible, asynchronous collaboration. The role blends subject matter expertise with rigorous evaluation to strengthen AI benchmarks in medicine and health care.