Head of Evals, AI Red Teaming
Trajectory Labs, PBC
Posted
Sep 10, 2026
Location
Remote
Type
Full-time
Compensation
$200000 - $400000
Mission
What you will drive
Lead evaluation quality and design for AI red-teaming safety tests deployed to frontier model developers.
- Review and approve red-teaming tasks, transcripts, and grades to ensure evals teach models robust defenses against prompt injection.
- Design new evaluation methodologies and environments targeting gaps in model robustness, informed by failure analysis.
- Build agent-based automation tools and checkers to scale your review judgment without proportional headcount growth.
- Requires calibrated eval judgment, fluency with LLMs and coding agents, and sustained attention to detail across hundreds of reviews.
- Desirable: prior experience evaluating benchmarks, building LLM judges, designing eval environments, or prompt injection and red-teaming work.
Profile
What makes you a great fit
Lead evaluation quality and design for AI red-teaming safety tests deployed to frontier model developers.
- Review and approve red-teaming tasks, transcripts, and grades to ensure evals teach models robust defenses against prompt injection.
- Design new evaluation methodologies and environments targeting gaps in model robustness, informed by failure analysis.
- Build agent-based automation tools and checkers to scale your review judgment without proportional headcount growth.
- Requires calibrated eval judgment, fluency with LLMs and coding agents, and sustained attention to detail across hundreds of reviews.
- Desirable: prior experience evaluating benchmarks, building LLM judges, designing eval environments, or prompt injection and red-teaming work.
About
Inside Trajectory Labs, PBC
Trajectory Labs is a startup that builds RL environments for frontier AI labs to train robust, secure, and reliable models.