Mercor
MercorVerified Source
Remote

Applied Psychology Benchmark Specialist | $50-$63/hr Remote

50–63/hr
Remote
Posted August 14, 2026
hourly
17 openings

Overview

A remote role for seasoned psychologists to author and review rigorous assessment items for an AI research effort. You’ll craft original multiple-choice questions and validate existing items to help define benchmarks that advance AI understanding of psychology. The work spans areas like psychometrics, digital health, and consumer psychology, with both creation and verification responsibilities. This opportunity blends scholarly rigor with flexible, asynchronous engagement and compensation on an hourly basis.

What You'll Do6

  • 1Create original psychology questions that probe deep theoretical understanding and avoid basic recall
  • 2Review and polish pre-written items for clarity, accuracy, and rigor, noting any edits
  • 3Assign difficulty levels such as Medium, Hard, or Expert to each question and justify the rating
  • 4Provide a correct answer plus 9 plausible distractors that challenge advanced solvers
  • 5Document step-by-step reasoning or solution approaches when required and include 1–5 scholarly references per item
  • 6Flag issues in question clarity or solvability during verification and explain amendments

Requirements5

  • 1PhD, PsyD, or doctoral candidate in Psychology or a closely related field
  • 2Strong mastery of graduate-level theory, research methods, and current empirical literature
  • 3Excellent written English with the ability to articulate complex ideas clearly
  • 4Experience with or exposure to psychometrics, assessment design, or related domains
  • 5Clinical licensure or relevant research publications are a plus

Who Should Apply

Ideal candidates are senior researchers or doctoral candidates who excel at translating complex psychology concepts into rigorous assessment material, comfortable working independently, and able to manage multiple items with careful attention to detail.

Salary Insight

Compensation ranges from $50.00 to $63.00 per hour, with remote, asynchronous engagement and an hourly arrangement.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

psychometricsassessment designgraduate-level psychologyresearch methodologyacademic writing

Application Tip

Submit a concise portfolio sample of prior assessment items you authored or reviewed, plus a brief note outlining how you would approach creating or vetting a benchmark item for AI research.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

4h agoRemotehourly

Applied Computer Science Benchmark Specialist | $66-$84/hr Remote

A remote, hourly role focused on building and reviewing high quality AI research benchmark content for computer science. You’ll craft and validate rigorous multiple choice questions across core CS topics, assess answer quality, and contribute to gold-standard datasets that push AI capabilities forward. Expect a split between creating new questions and verifying existing ones, with opportunities to shape benchmarks used by researchers worldwide.

66–84/hr
· 27 openings
phd or doctoral candidatecomputer science theoryalgorithms+3 more
Mercor

Mercor

4h agoRemotehourly

Applied Engineering Benchmark Specialist | $61-$77/hr Remote

The role focuses on creating and validating rigorous academic assessment content for an AI research effort. You will author and review multiple-choice questions across core engineering topics, judge solution quality, and help build gold-standard benchmarks to push AI capabilities forward. This is a remote, hourly engagement, with tasks that blend technical evaluation and high-quality writing using a clear, structured approach.

61–77/hr
· 26 openings
semiconductor design & manufacturingcontrol sciencemechatronics+5 more
Mercor

Mercor

4h agoRemotehourly

Applied Philosophy Benchmark Specialist | $50-$63/hr Remote

A remote role focused on creating and vetting high quality philosophy assessment content for an AI research initiative. You will craft and critique multiple choice questions across core philosophy domains, while helping set gold standard benchmarks for advancing AI capabilities. The work blends deep philosophical expertise with rigorous evaluation, and offers flexible, asynchronous engagement.

50–63/hr
· 13 openings
philosophyformal logicacademic writing+1 more
Mercor

Mercor

4h agoRemotehourly

Applied Physics Benchmark Specialist | $61-$77/hr Remote

We are looking for seasoned physicists to craft and review advanced academic assessment content for an AI research project. You’ll develop and validate high quality multiple-choice questions across key physics areas, assess solution quality, and help establish benchmark materials that push AI capabilities forward. The role is remote and hourly, with flexible, asynchronous collaboration. You’ll work in two modes: authoring original questions and verifying pre-written items for accuracy and rigor.

61–77/hr
· 19 openings
phd in physicsgraduate-level physics conceptsproblem design+4 more
Mercor

Mercor

4h agoRemotehourly

Applied History & Political Science Benchmark Specialist | $44-$56/hr Remote

This remote role focuses on creating and validating academic assessment content for an AI research initiative. You’ll craft and review high quality multiple choice items in history and political science, help determine gold standard benchmarks, and contribute to improving AI understanding in these fields. The work blends scholarly rigor with practical QA to support advanced AI systems. Expect a flexible, asynchronous arrangement that taps deep expertise in your domains.

44–56/hr
· 17 openings
question authoringquestion verificationhistoriography+2 more
Mercor

Mercor

4h agoRemotehourly

Applied Biology Benchmark Specialist | $60-$75/hr Remote

One sentence on its own. We’re looking for seasoned biologists to craft and review high quality academic assessment content for an AI research project. You’ll develop and verify rigorous biology multiple-choice items, assess solution quality, and help set gold-standard benchmarks that push AI capabilities forward. The role spans authoring original questions and validating existing ones across several biology domains using a remote, asynchronous setup and competitive hourly pay.

60–75/hr
· 17 openings
biologymolecular biologybiochemistry+4 more