Mercor
MercorVerified Source
Remote

Applied Engineering Benchmark Specialist | $61-$77/hr Remote

61–77/hr
Remote
Posted August 14, 2026
hourly
26 openings

Overview

The role focuses on creating and validating rigorous academic assessment content for an AI research effort. You will author and review multiple-choice questions across core engineering topics, judge solution quality, and help build gold-standard benchmarks to push AI capabilities forward. This is a remote, hourly engagement, with tasks that blend technical evaluation and high-quality writing using a clear, structured approach.

What You'll Do5

  • 1Craft original engineering questions that probe deep understanding rather than mere memorization, and assign a difficulty level
  • 2Review existing questions for clarity, completeness, and technical accuracy, making precise edits
  • 3Evaluate each item’s challenge level as Medium, Hard, or Expert, and provide a correct answer plus 9 convincing distractors
  • 4Document step-by-step reasoning solutions in a concise, readable format
  • 5Attach 1 to 5 scholarly references per item from reputable sources and flag any issues that affect solvability or clarity

Requirements5

  • 1Doctoral candidate or PhD holder in Engineering or a closely related field, or a strong demonstration of depth in a subdomain
  • 2Master’s degree with exceptional depth may be considered
  • 3Solid grasp of graduate-level engineering concepts, applied math, and relevant standards
  • 4Licensure as a Professional Engineer or substantial industry experience is a plus
  • 5Excellent written English with the ability to convey complex ideas clearly and concisely

Who Should Apply

We’re looking for engineers who enjoy rigorous problem-solving, clear technical writing, and contributing to benchmarks that drive AI research. Ideal candidates combine deep domain knowledge with a knack for precise communication and a bias for high-quality exam content.

Salary Insight

Hourly compensation ranges from $61.00 to $77.00, with remote, asynchronous engagement.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

semiconductor design & manufacturingcontrol sciencemechatronicsbioinstrumentation & biotechnologyapplied mathematicsacademic writingquestion designtechnical editing

Application Tip

Share a concise portfolio snippet showing a sample question and a brief justification of its difficulty to demonstrate your ability to craft and vet rigorous content.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

4h agoRemotehourly

Applied Computer Science Benchmark Specialist | $66-$84/hr Remote

A remote, hourly role focused on building and reviewing high quality AI research benchmark content for computer science. You’ll craft and validate rigorous multiple choice questions across core CS topics, assess answer quality, and contribute to gold-standard datasets that push AI capabilities forward. Expect a split between creating new questions and verifying existing ones, with opportunities to shape benchmarks used by researchers worldwide.

66–84/hr
· 27 openings
phd or doctoral candidatecomputer science theoryalgorithms+3 more
Mercor

Mercor

4h agoRemotehourly

Applied Mathematics Benchmark Specialist | $61-$77/hr Remote

A remote role focused on creating and validating rigorous math assessment content for an AI research initiative. You’ll write and review advanced multiple choice questions across key mathematics domains, assess solution quality, and help set benchmarks that push AI capabilities forward. The work blends mathematical depth with clear, instructional problem design and documentation. This position stands out for its blend of academic rigor and practical benchmarking in AI research.

61–77/hr
· 19 openings
question designformal proof writingacademic referencing+2 more
Mercor

Mercor

4h agoRemotehourly

Applied Physics Benchmark Specialist | $61-$77/hr Remote

We are looking for seasoned physicists to craft and review advanced academic assessment content for an AI research project. You’ll develop and validate high quality multiple-choice questions across key physics areas, assess solution quality, and help establish benchmark materials that push AI capabilities forward. The role is remote and hourly, with flexible, asynchronous collaboration. You’ll work in two modes: authoring original questions and verifying pre-written items for accuracy and rigor.

61–77/hr
· 19 openings
phd in physicsgraduate-level physics conceptsproblem design+4 more
Mercor

Mercor

4h agoRemotehourly

Applied Biology Benchmark Specialist | $60-$75/hr Remote

One sentence on its own. We’re looking for seasoned biologists to craft and review high quality academic assessment content for an AI research project. You’ll develop and verify rigorous biology multiple-choice items, assess solution quality, and help set gold-standard benchmarks that push AI capabilities forward. The role spans authoring original questions and validating existing ones across several biology domains using a remote, asynchronous setup and competitive hourly pay.

60–75/hr
· 17 openings
biologymolecular biologybiochemistry+4 more
Mercor

Mercor

4h agoRemotehourly

Applied Chemistry Benchmark Specialist | $61-$77/hr Remote

We’re looking for seasoned chemists to craft and review high quality academic assessment content for an AI research initiative. In this role you’ll develop and validate rigorous chemistry multiple choice questions, evaluate answer quality, and help set gold-standard benchmarks that push AI understanding forward. Work is remote and project-based, with flexible hours. You’ll work in areas spanning materials, polymer and electronic chemistry, industrial processes, energy storage, environmental chemistry, pharmaceutical and agrochemical chemistry, as well as consumer and food chemistry.

61–77/hr
· 19 openings
phd or doctoral candidacychemistry knowledge at graduate levelproblem design and/or olympiad writing experience+1 more
Mercor

Mercor

4h agoRemotehourly

Applied Philosophy Benchmark Specialist | $50-$63/hr Remote

A remote role focused on creating and vetting high quality philosophy assessment content for an AI research initiative. You will craft and critique multiple choice questions across core philosophy domains, while helping set gold standard benchmarks for advancing AI capabilities. The work blends deep philosophical expertise with rigorous evaluation, and offers flexible, asynchronous engagement.

50–63/hr
· 13 openings
philosophyformal logicacademic writing+1 more