Mercor
MercorVerified Source
Remote

AI Safety Practitioner | $60-$70/hr Remote

60–70/hr
Remote
Posted July 23, 2026
hourly
4 openings

Overview

AIUC is looking for experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models on complex, policy-sensitive topics. In this role, you'll review AI-generated responses for factual accuracy, policy compliance, and overall quality, helping improve how these models handle ambiguous and high-risk grey areas. You'll apply structured evaluation rubrics and provide feedback that directly shapes model behavior. This is a fully remote, hourly freelance position that's ideal for professionals from journalism, policy, or scientific backgrounds who want to apply their expertise to AI alignment.

What You'll Do6

  • 1Review AI-generated outputs for safety, accuracy, and alignment with established policies and quality standards.
  • 2Assess content across sensitive domains including misinformation, political persuasion, self-harm, violence, cyber threats, and biosecurity.
  • 3Apply and help refine evaluation prompts and rubrics used for RLHF, SFT, and AI safety benchmarking.
  • 4Flag unsafe outputs, hallucinations, reasoning errors, and policy violations, then document your findings clearly.
  • 5Provide structured, actionable feedback that helps researchers improve model alignment and safety performance.
  • 6Work alongside AI researchers and safety teams on ongoing evaluation projects and contribute to iterative improvements.

Requirements5

  • 1A bachelor's degree or higher in a relevant field such as journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science, or a related discipline.
  • 2At least 5 years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, or a comparable field.
  • 3Strong written English, critical thinking, and analytical reasoning skills, with the ability to evaluate nuanced and policy-sensitive scenarios consistently.
  • 4Familiarity with AI safety concepts, RLHF, SFT, or content moderation is a plus, as is experience developing evaluation rubrics.
  • 5Experience reviewing complex, high-risk, or ambiguous content is preferred.

Who Should Apply

You're the kind of person who thrives on nuance and can make careful judgments in ambiguous situations. You have a sharp analytical eye, excellent writing skills, and a genuine interest in making AI systems safer and more reliable. Whether you come from journalism, policy analysis, psychology, or a technical field, you're excited to apply your expertise to real-world AI evaluation challenges. You're comfortable working independently in a remote environment and can deliver consistent, high-quality assessments.

Salary Insight

This is an hourly remote freelance position paying $60.00 - $70.00 per hour.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

ai safetyrlhfsfttrust & safetycontent moderationevaluation rubricsmisinformationbiosecuritycyber securitypolicy analysiscritical reasoningwritten communicationjournalismpsychologypublic policycomputer science

Application Tip

Tailor your application to highlight your experience with ambiguous, policy-sensitive content—use specific examples from past work that show your ability to analyze complex scenarios and articulate clear reasoning. Mention any familiarity with RLHF, SFT, or content moderation to stand out.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

18d agoRemotehourly

AI Safety Red Teamer | $70-$84/hr Remote

This role focuses on stress-testing some of the most advanced AI systems in the world. As an AI Safety Red Teamer, you'll design tricky prompts, hunt for vulnerabilities, and push models to see how they handle dangerous or ambiguous topics. You'll work fully remotely, collaborating with researchers who care deeply about alignment and safety. If you enjoy breaking things to make them stronger, this is a great fit.

70–84/hr
· 4 openings
ai safetyred teamingadversarial testing+13 more

Mercor

10h agoRemotecontract

AI Safety Red Teamer Mercor

Mercor seeks an AI Safety Red Teamer to conduct adversarial evaluations of frontier AI models. This fully remote contract role offers competitive compensation up to $84/hour. You will design adversarial prompts identify jailbreaks evaluate model robustness and document vulnerabilities. The ideal candidate has strong analytical reasoning and experience in AI safety red teaming.

70K–84K
Adversarial prompt designAI safety/red teamingAnalytical reasoning+5 more

YO AI Labs

13h agoRemotecontract

AI Data Science Expert, Model Evaluation & Prompt Engineering

As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.

Competitive salary
data sciencemachine learningprompt engineering+2 more
Mercor

Mercor

18d agoRemotehourly

Child & Adolescent Mental Health Clinical Advisor (AI Safety Benchmark Project) | $80-$150/hr Remote

This remote role puts your clinical expertise to work on an AI safety project: you'll help build a clinician-informed benchmark for evaluating how AI companion chatbots respond to teens in mental health crises. The goal is to give AI developers and researchers a rigorous, real-world grounding for handling sensitive issues like suicide risk, therapeutic boundaries, and emotional dependency. You'll author realistic case scenarios, review AI-generated conversations, and shape evaluation rubrics alongside a leading AI research organisation. It's a unique chance to influence the responsible deployment of conversational AI in adolescent mental health.

80–150/hr
· 4 openings
child and adolescent psychiatryclinical psychologysuicide prevention+8 more
Mercor

Mercor

11d agoRemotehourly

AI Safety Experts — English & Indonesian | $17-$25/hr Remote

We're hiring bilingual AI safety experts to stress-test conversational AI systems. In this remote contract role, you'll probe AI models for vulnerabilities like jailbreaks, prompt injections, and bias exploits, producing data that makes AI safer. You'll work with a red team that simulates real-world adversarial attacks, using structured playbooks and taxonomies. What sets this apart: you'll shape the safety of frontier AI products while earning $17–$25 per hour.

17–25/hr
· 10 openings
prompt injectionjailbreakred teaming+11 more
Mercor

Mercor

1d agoRemotehourly

AI Safety Experts — English & Finnish | $48-$62/hr Remote

Mercor is assembling a hand-picked red team of bilingual AI safety experts to stress-test conversational AI models in English and Finnish. This remote, hourly role focuses on probing models for vulnerabilities like jailbreaks, prompt injections, and bias exploits, then turning those discoveries into structured data that helps customers harden their AI systems. You'll follow established taxonomies and playbooks, with the option to skip higher-sensitivity projects. It's a unique chance to apply adversarial thinking at the frontier of AI safety while earning a competitive hourly rate.

48–62/hr
· 20 openings
prompt injectionjailbreakadversarial ml+14 more