Safeguards Enforcement Lead, User Well-Being

Anthropic

Location
Remote (US)
Type
Full-time
Compensation
$285,000 – $330,000 per year
Posted
Sep 02, 2026

Confirmed open on Anthropic's own job board on Oct 9, 2026. We re-check every day.

Job description

Lead enforcement operations for child safety, mental health, and abuse prevention across AI systems, managing review teams and detection workflows.
- Manage content review teams across child safety, mental health, abuse, and age assurance policy areas; oversee quality assurance and escalation processes.
- Design and scale enforcement workflows while partnering with Engineering and Data Science to optimize detection models and automated systems.
- Develop enforcement guidelines, monitor emerging misuse trends, and coordinate external reporting obligations to child safety bodies.
- Requires experience managing trust and safety or content moderation operations with direct exposure to child safety and exploitation harms.
- Requires SQL or data analysis proficiency and ability to identify emerging risks and communicate findings to cross-functional teams.

About Anthropic

Website

Anthropic is a frontier AI research and product company, with teams working on alignment, policy, and security. We post specific opportunities at Anthropic that we think may be high impact. We do not necessarily recommend working at other positions at Anthropic. You can read concerns about doing harm by working at a frontier AI company in our career review on the topic.

Share

Tweet Share WhatsApp Email Facebook

Cover letter and interview prep

Draft a cover letter for this role or practise the questions you are likely to be asked.

Open application tools

Want more roles like this?

Describe what you want and get your strongest matching jobs by email.

Find matching jobs

Related searches

Safeguards Enforcement Lead, User Well-Being

Anthropic

Apply on company site