Strategic Projects Lead, AI Safety & Red Teaming
Overview
You will own the delivery of high-stakes data programs for frontier AI labs, shaping how AI Safety and alignment work gets done at scale. Working directly with safety, alignment, and trust & safety teams at labs like OpenAI, Anthropic, and Google DeepMind, you will translate their hardest data problems into executable projects. You will lead specialized red teamers and domain experts, build repeatable methodologies for adversarial testing, and drive revenue from first conversation to final delivery. This is a founding-level role at a fast-growing YC-backed company with $100M+ revenue run rate and a $300M valuation.
What You'll Do6
- 1Partner with safety, alignment, and trust & safety teams at frontier AI labs to scope their hardest data problems and translate them into deliverable programs that drive revenue.
- 2Recruit and lead specialized contributors: red teamers, security researchers, trust & safety practitioners, and domain risk experts, holding a high bar on output quality.
- 3Design specifications for adversarial datasets, jailbreak and attack taxonomies, refusal-boundary sets, and safety benchmarks that labs use to measure model behavior.
- 4Build repeatable red teaming machinery: human red team campaigns, automated attack generation, and coverage tracking against harm taxonomies, avoiding one-off deliverables.
- 5Own projects end-to-end from first conversation through final delivery, in a fast-moving environment where the spec often does not exist yet.
- 6Support initiatives across building, analysis, coordination, and execution as the safety practice scales.
Requirements8
- 12+ years in AI safety, red teaming, trust & safety, adversarial ML, or security research at a frontier AI lab, FAANG, top security firm, or equivalent.
- 2Hands-on adversarial experience: jailbreaking or stress-testing frontier models, prompt injection research, offensive security, penetration testing, bug bounty, CTFs, or trust & safety investigations.
- 3Fluency with harm taxonomies, evaluation design, safety frameworks, and model policy.
- 4High agency and ability to execute in ambiguity, defining problems before solving them.
- 5Strong writing skills to specify risk precisely enough for 50+ people to execute against.
- 6Genuine passion for AI and entrepreneurship, with a track record of competitive success.
- 7Strong leadership and communication skills.
- 8Python proficiency and the ability to write production-quality code.
Salary Insight
Salary not disclosed in listing
Location
Required Skills
Similar open positions
Explore active roles that match your skills and interests.

Technogen, Inc.
VerifiedAI Safety & Responsible AI Lead
You will own the AI safety strategy for a Fortune 50 client, ensuring responsible AI practices across machine learning pipelines and model deployments. You will lead a team of 5 engineers and data scientists, working with Python, TensorFlow, and PyTorch. You will define governance frameworks, conduct risk assessments, and implement safety guardrails for AI systems that serve millions of users. This role stands out for its direct influence on enterprise-wide AI ethics and regulatory compliance.

Mercor
VerifiedAI Safety Red Teamer | $70-$84/hr Remote
This role focuses on stress-testing some of the most advanced AI systems in the world. As an AI Safety Red Teamer, you'll design tricky prompts, hunt for vulnerabilities, and push models to see how they handle dangerous or ambiguous topics. You'll work fully remotely, collaborating with researchers who care deeply about alignment and safety. If you enjoy breaking things to make them stronger, this is a great fit.
Afterquery
VerifiedStrategic Projects Lead, Legal AI Data
Own legal data and evaluation programs for frontier AI labs at AfterQuery, an applied research lab serving every major AI developer. You will define what correct legal reasoning looks like, scoping problems, translating them into deliverable programs, and driving revenue. Report directly to the founding team and collaborate with engineers and researchers building the infrastructure that powers foundation model training. This role offers founding-level impact, meaningful equity, and a front-row seat to the defining moment in AI.
Mercor
VerifiedAI Safety Red Teamer Mercor
Mercor seeks an AI Safety Red Teamer to conduct adversarial evaluations of frontier AI models. This fully remote contract role offers competitive compensation up to $84/hour. You will design adversarial prompts identify jailbreaks evaluate model robustness and document vulnerabilities. The ideal candidate has strong analytical reasoning and experience in AI safety red teaming.
Mercor
VerifiedAI Safety Expert, Red Team & Adversarial ML
As an AI Safety Expert on the red team, you will probe conversational AI models and agents to uncover jailbreaks, prompt injections, and bias exploits at scale. You will join Mercor, a San Francisco-based talent network backed by Benchmark, General Catalyst, and Peter Thiel, working remotely with a team of elite technical and creative professionals. Your findings will directly shape safer AI deployments for leading research labs. This contract role offers $48–$62/hour and requires fluency in English and Finnish.
Clera
VerifiedFounding Engineer, AI Agents & Data Infrastructure
Join a three-person founding team in New York and build the full-stack systems powering agentic operations that handle thousands of jobs daily. Own work across frontend, data infrastructure, and AI orchestration, making AI agents dependable through reliable tooling, durable execution, and strong observability. This is a high-ownership, end-to-end role at a Y Combinator–backed company in financial research and healthcare data, where you will inspect data, ship improvements, and fix production issues. What sets this role apart: you will add a new information source on day one and ship a production module within 30 days.