AI Safety & Governance Full-time

Head of Evals, AI Red Teaming

Trajectory Labs, PBC

Posted

Sep 10, 2026

Location

Remote

Type

Full-time

Compensation

$200000 - $400000

Mission

What you will drive

Lead evaluation quality and design for AI red-teaming safety tests deployed to frontier model developers.
- Review and approve red-teaming tasks, transcripts, and grades to ensure evals teach models robust defenses against prompt injection.
- Design new evaluation methodologies and environments targeting gaps in model robustness, informed by failure analysis.
- Build agent-based automation tools and checkers to scale your review judgment without proportional headcount growth.
- Requires calibrated eval judgment, fluency with LLMs and coding agents, and sustained attention to detail across hundreds of reviews.
- Desirable: prior experience evaluating benchmarks, building LLM judges, designing eval environments, or prompt injection and red-teaming work.

Profile

What makes you a great fit

Lead evaluation quality and design for AI red-teaming safety tests deployed to frontier model developers.
- Review and approve red-teaming tasks, transcripts, and grades to ensure evals teach models robust defenses against prompt injection.
- Design new evaluation methodologies and environments targeting gaps in model robustness, informed by failure analysis.
- Build agent-based automation tools and checkers to scale your review judgment without proportional headcount growth.
- Requires calibrated eval judgment, fluency with LLMs and coding agents, and sustained attention to detail across hundreds of reviews.
- Desirable: prior experience evaluating benchmarks, building LLM judges, designing eval environments, or prompt injection and red-teaming work.

About

Inside Trajectory Labs, PBC

Visit site →

Trajectory Labs is a startup that builds RL environments for frontier AI labs to train robust, secure, and reliable models.