
Data Science Expert | $120-$170/hr Remote
Overview
This is a remote, hourly contract role for experienced data scientists who want to shape how AI systems are evaluated. You won't be building models or dashboards yourself — instead, you'll define what great data science work looks like and score AI-generated output against your criteria. The work sits at the intersection of data science and AI evaluation, making it a great fit for someone who enjoys judgment calls and rigorous reasoning. You'll collaborate with a leading AI research organization through Mercor, with pay at $120–$170 per hour.
What You'll Do5
- 1Develop clear, task-specific rubrics for assessing real-world data science deliverables like analyses, predictive models, dashboards, experiment summaries, and written recommendations.
- 2Evaluate both AI-generated and human-produced work samples against those rubrics, providing detailed written rationale for every score you assign.
- 3Apply consistent, evidence-based judgment so your scoring is reproducible and can withstand scrutiny from other reviewers.
- 4Act on structured feedback from senior reviewers and refine your criteria and scoring approach quickly.
- 5Help calibrate quality standards by participating in peer review and alignment discussions.
Requirements5
- 15+ years of hands-on data science experience in an industry setting, ideally at a top-tier tech company in business, product, or growth analytics.
- 2Strong command of experiment design, A/B testing, metric definition, and SQL/Python for analysis and communication.
- 3Exceptional written communication skills — you can explain complex reasoning clearly and concisely.
- 4Highly detail-oriented and consistent, comfortable having your judgment reviewed and calibrated against peers.
- 5Bonus: prior experience with AI training, evaluation, or human-data labeling projects.
Who Should Apply
You’re a seasoned data scientist who’s excelled in fast-paced product or growth environments and can articulate what makes analytics work good or bad. You enjoy the precision of writing evaluation criteria and defending your scores with evidence. If you’ve ever been the person others come to for a second opinion on an analysis, this role will feel natural.
Salary Insight
Hourly rate of $120–$170 depending on experience and qualifications.
Location
Required Skills
Application Tip
Highlight a past project where you had to define quality metrics or evaluate other data scientists' work — for example, a rubric you created for a hackathon or a peer-review process you led. Concrete examples of your written evaluation style will strengthen your application.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedData Science and Analytics Experts | $60-$70/hr Remote
This remote, hourly role puts your senior data science expertise to work building evaluation tasks that test AI systems against the realities of Fortune 500 enterprise data operations. You'll design realistic scenarios, craft reference outputs, and author scoring rubrics that capture how top analytics leaders think. If you've lived the complexity of enterprise pipelines, governance, and model deployment, this is a chance to shape AI evaluation at the highest level.
Mercor
VerifiedData Evaluator, AI Quality & Tech Assessment
You will evaluate AI-generated artifacts against domain-specific quality rubrics, owning accuracy and presentation standards for documents, spreadsheets, and slide decks. Reporting to a decentralized team at Mercor, you'll work independently asynchronously with leading AI research labs. With 5+ years in Software, AI, or IT, you'll apply deep expertise to grade outputs with rigor. Contract pays $80–$120/hour remotely.
YO AI Labs
VerifiedAI Data Science Expert, Model Evaluation & Prompt Engineering
As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.

Mercor
VerifiedData Science & Quantitative Analysis Expert | $60-$90/hr Remote
A premier AI research organization is assembling a team of quantitative experts to help design the next generation of evaluation benchmarks for frontier models. In this role, you'll create realistic, hands-on data analysis challenges that push models to their limits — cleaning messy datasets, comparing statistical methods, and verifying whether model outputs hold up to scrutiny. This is a fully remote, W-2 position with Cincinnatus LLC, working about 35 hours per week and collaborating closely with the lab's researchers. If you're passionate about the science behind AI and want your analytical expertise to directly shape how models are measured, this is a unique opportunity.
Mercor
VerifiedB2B Sales Expert - Evaluator
Mercor matches elite creative and technical talent with top AI research labs. Based in San Francisco we evaluate AI researchers. Remote contract position.

Mercor
VerifiedData Scientist Talent Network | $100-$150/hr Remote
This is a remote talent network for experienced data scientists interested in evaluating AI systems' real-world data science performance for leading AI research organizations. There's no active project right now — you're joining a pool of vetted experts who may be contacted as relevant opportunities arise. If you have a strong background in Python, SQL, and statistical modeling, and you enjoy judging complex analytical work, this is a chance to apply that expertise on high-impact AI evaluation projects.