
AI Safety Practitioner | $60-$70/hr Remote
Overview
AIUC is looking for experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models on complex, policy-sensitive topics. In this role, you'll review AI-generated responses for factual accuracy, policy compliance, and overall quality, helping improve how these models handle ambiguous and high-risk grey areas. You'll apply structured evaluation rubrics and provide feedback that directly shapes model behavior. This is a fully remote, hourly freelance position that's ideal for professionals from journalism, policy, or scientific backgrounds who want to apply their expertise to AI alignment.
What You'll Do6
- 1Review AI-generated outputs for safety, accuracy, and alignment with established policies and quality standards.
- 2Assess content across sensitive domains including misinformation, political persuasion, self-harm, violence, cyber threats, and biosecurity.
- 3Apply and help refine evaluation prompts and rubrics used for RLHF, SFT, and AI safety benchmarking.
- 4Flag unsafe outputs, hallucinations, reasoning errors, and policy violations, then document your findings clearly.
- 5Provide structured, actionable feedback that helps researchers improve model alignment and safety performance.
- 6Work alongside AI researchers and safety teams on ongoing evaluation projects and contribute to iterative improvements.
Requirements5
- 1A bachelor's degree or higher in a relevant field such as journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science, or a related discipline.
- 2At least 5 years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, or a comparable field.
- 3Strong written English, critical thinking, and analytical reasoning skills, with the ability to evaluate nuanced and policy-sensitive scenarios consistently.
- 4Familiarity with AI safety concepts, RLHF, SFT, or content moderation is a plus, as is experience developing evaluation rubrics.
- 5Experience reviewing complex, high-risk, or ambiguous content is preferred.
Who Should Apply
You're the kind of person who thrives on nuance and can make careful judgments in ambiguous situations. You have a sharp analytical eye, excellent writing skills, and a genuine interest in making AI systems safer and more reliable. Whether you come from journalism, policy analysis, psychology, or a technical field, you're excited to apply your expertise to real-world AI evaluation challenges. You're comfortable working independently in a remote environment and can deliver consistent, high-quality assessments.
Salary Insight
This is an hourly remote freelance position paying $60.00 - $70.00 per hour.
Location
Required Skills
Application Tip
Tailor your application to highlight your experience with ambiguous, policy-sensitive content—use specific examples from past work that show your ability to analyze complex scenarios and articulate clear reasoning. Mention any familiarity with RLHF, SFT, or content moderation to stand out.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedAI Safety Red Teamer | $70-$84/hr Remote
This role focuses on stress-testing some of the most advanced AI systems in the world. As an AI Safety Red Teamer, you'll design tricky prompts, hunt for vulnerabilities, and push models to see how they handle dangerous or ambiguous topics. You'll work fully remotely, collaborating with researchers who care deeply about alignment and safety. If you enjoy breaking things to make them stronger, this is a great fit.
Mercor
VerifiedAI Safety Red Teamer Mercor
Mercor seeks an AI Safety Red Teamer to conduct adversarial evaluations of frontier AI models. This fully remote contract role offers competitive compensation up to $84/hour. You will design adversarial prompts identify jailbreaks evaluate model robustness and document vulnerabilities. The ideal candidate has strong analytical reasoning and experience in AI safety red teaming.
YO AI Labs
VerifiedAI Data Science Expert, Model Evaluation & Prompt Engineering
As an AI Data Science Domain Expert, you own the evaluation and refinement of AI-generated technical content, directly shaping the reasoning and accuracy of next-generation AI systems. You work remotely with a cross-functional team of data scientists, engineers, and product leads, delivering rubric-based assessments and structured feedback that drive model improvements. No prior AI experience is required; your expertise in data science, analytical thinking, and communication is what matters most. This contract role offers the chance to contribute to cutting-edge model development while honing skills in prompt engineering and RLHF.

Mercor
VerifiedChild & Adolescent Mental Health Clinical Advisor (AI Safety Benchmark Project) | $80-$150/hr Remote
This remote role puts your clinical expertise to work on an AI safety project: you'll help build a clinician-informed benchmark for evaluating how AI companion chatbots respond to teens in mental health crises. The goal is to give AI developers and researchers a rigorous, real-world grounding for handling sensitive issues like suicide risk, therapeutic boundaries, and emotional dependency. You'll author realistic case scenarios, review AI-generated conversations, and shape evaluation rubrics alongside a leading AI research organisation. It's a unique chance to influence the responsible deployment of conversational AI in adolescent mental health.

Mercor
VerifiedAI Safety Experts — English & Indonesian | $17-$25/hr Remote
We're hiring bilingual AI safety experts to stress-test conversational AI systems. In this remote contract role, you'll probe AI models for vulnerabilities like jailbreaks, prompt injections, and bias exploits, producing data that makes AI safer. You'll work with a red team that simulates real-world adversarial attacks, using structured playbooks and taxonomies. What sets this apart: you'll shape the safety of frontier AI products while earning $17–$25 per hour.

Mercor
VerifiedAI Safety Experts — English & Finnish | $48-$62/hr Remote
Mercor is assembling a hand-picked red team of bilingual AI safety experts to stress-test conversational AI models in English and Finnish. This remote, hourly role focuses on probing models for vulnerabilities like jailbreaks, prompt injections, and bias exploits, then turning those discoveries into structured data that helps customers harden their AI systems. You'll follow established taxonomies and playbooks, with the option to skip higher-sensitivity projects. It's a unique chance to apply adversarial thinking at the frontier of AI safety while earning a competitive hourly rate.