AfterqueryVerified Source

Strategic Projects Lead, AI Safety & Red Teaming

Onsite · San Francisco, California
Posted August 13, 2026
payroll

Overview

You will own the delivery of high-stakes data programs for frontier AI labs, shaping how AI Safety and alignment work gets done at scale. Working directly with safety, alignment, and trust & safety teams at labs like OpenAI, Anthropic, and Google DeepMind, you will translate their hardest data problems into executable projects. You will lead specialized red teamers and domain experts, build repeatable methodologies for adversarial testing, and drive revenue from first conversation to final delivery. This is a founding-level role at a fast-growing YC-backed company with $100M+ revenue run rate and a $300M valuation.

What You'll Do6

  • 1Partner with safety, alignment, and trust & safety teams at frontier AI labs to scope their hardest data problems and translate them into deliverable programs that drive revenue.
  • 2Recruit and lead specialized contributors: red teamers, security researchers, trust & safety practitioners, and domain risk experts, holding a high bar on output quality.
  • 3Design specifications for adversarial datasets, jailbreak and attack taxonomies, refusal-boundary sets, and safety benchmarks that labs use to measure model behavior.
  • 4Build repeatable red teaming machinery: human red team campaigns, automated attack generation, and coverage tracking against harm taxonomies, avoiding one-off deliverables.
  • 5Own projects end-to-end from first conversation through final delivery, in a fast-moving environment where the spec often does not exist yet.
  • 6Support initiatives across building, analysis, coordination, and execution as the safety practice scales.

Requirements8

  • 12+ years in AI safety, red teaming, trust & safety, adversarial ML, or security research at a frontier AI lab, FAANG, top security firm, or equivalent.
  • 2Hands-on adversarial experience: jailbreaking or stress-testing frontier models, prompt injection research, offensive security, penetration testing, bug bounty, CTFs, or trust & safety investigations.
  • 3Fluency with harm taxonomies, evaluation design, safety frameworks, and model policy.
  • 4High agency and ability to execute in ambiguity, defining problems before solving them.
  • 5Strong writing skills to specify risk precisely enough for 50+ people to execute against.
  • 6Genuine passion for AI and entrepreneurship, with a track record of competitive success.
  • 7Strong leadership and communication skills.
  • 8Python proficiency and the ability to write production-quality code.

Salary Insight

Salary not disclosed in listing

Location

Typeonsite
LocationSan Francisco, California

Required Skills

pythonred teamingai safetysecurity researchprompt injection
Share:

Similar open positions

Explore active roles that match your skills and interests.

Technogen, Inc.

Technogen, Inc.

12h agoNew York, New Yorkpayroll

AI Safety & Responsible AI Lead

You will own the AI safety strategy for a Fortune 50 client, ensuring responsible AI practices across machine learning pipelines and model deployments. You will lead a team of 5 engineers and data scientists, working with Python, TensorFlow, and PyTorch. You will define governance frameworks, conduct risk assessments, and implement safety guardrails for AI systems that serve millions of users. This role stands out for its direct influence on enterprise-wide AI ethics and regulatory compliance.

Competitive salary
pythontensorflowpytorch+2 more
Mercor

Mercor

19d agoRemotehourly

AI Safety Red Teamer | $70-$84/hr Remote

This role focuses on stress-testing some of the most advanced AI systems in the world. As an AI Safety Red Teamer, you'll design tricky prompts, hunt for vulnerabilities, and push models to see how they handle dangerous or ambiguous topics. You'll work fully remotely, collaborating with researchers who care deeply about alignment and safety. If you enjoy breaking things to make them stronger, this is a great fit.

70–84/hr
· 4 openings
ai safetyred teamingadversarial testing+13 more

Afterquery

12h agoSan Francisco, Californiapayroll

Strategic Projects Lead, Legal AI Data

Own legal data and evaluation programs for frontier AI labs at AfterQuery, an applied research lab serving every major AI developer. You will define what correct legal reasoning looks like, scoping problems, translating them into deliverable programs, and driving revenue. Report directly to the founding team and collaborate with engineers and researchers building the infrastructure that powers foundation model training. This role offers founding-level impact, meaningful equity, and a front-row seat to the defining moment in AI.

Competitive salary
pythonlegal reasoningdata analysis+2 more

Mercor

1d agoRemotecontract

AI Safety Red Teamer Mercor

Mercor seeks an AI Safety Red Teamer to conduct adversarial evaluations of frontier AI models. This fully remote contract role offers competitive compensation up to $84/hour. You will design adversarial prompts identify jailbreaks evaluate model robustness and document vulnerabilities. The ideal candidate has strong analytical reasoning and experience in AI safety red teaming.

70K–84K
Adversarial prompt designAI safety/red teamingAnalytical reasoning+5 more

Mercor

1d agoRemotepayroll

AI Safety Expert, Red Team & Adversarial ML

As an AI Safety Expert on the red team, you will probe conversational AI models and agents to uncover jailbreaks, prompt injections, and bias exploits at scale. You will join Mercor, a San Francisco-based talent network backed by Benchmark, General Catalyst, and Peter Thiel, working remotely with a team of elite technical and creative professionals. Your findings will directly shape safer AI deployments for leading research labs. This contract role offers $48–$62/hour and requires fluency in English and Finnish.

Competitive salary
red teamingadversarial mlcybersecurity+2 more

Clera

1d agoRemotepayroll

Founding Engineer, AI Agents & Data Infrastructure

Join a three-person founding team in New York and build the full-stack systems powering agentic operations that handle thousands of jobs daily. Own work across frontend, data infrastructure, and AI orchestration, making AI agents dependable through reliable tooling, durable execution, and strong observability. This is a high-ownership, end-to-end role at a Y Combinator–backed company in financial research and healthcare data, where you will inspect data, ship improvements, and fix production issues. What sets this role apart: you will add a new information source on day one and ship a production module within 30 days.

Competitive salary
pythonfastapitypescript+2 more