AI Safety & Governance Full-time

Fellows, AI Safety and Security

Anthropic

Posted

Aug 18, 2026

Location

Remote (US)

Type

Full-time

Compensation

Up to $223600

Mission

What you will drive

  • In this fellowship, you'll conduct a 4-month full-time empirical research project on AI safety or security aligned with Anthropic's priorities.
  • Work under direct mentorship from Anthropic researchers in areas such as scalable oversight or mechanistic interpretability.
  • Produce a public research output such as a paper submission from your empirical project.
  • Develop and test approaches to technical problems in AI safety or security using external resources.
  • Iterate rapidly on your research ideas based on mentor feedback and research outcomes.

Profile

What makes you a great fit

  • In this fellowship, you'll conduct a 4-month full-time empirical research project on AI safety or security aligned with Anthropic's priorities.
  • Work under direct mentorship from Anthropic researchers in areas such as scalable oversight or mechanistic interpretability.
  • Produce a public research output such as a paper submission from your empirical project.
  • Develop and test approaches to technical problems in AI safety or security using external resources.
  • Iterate rapidly on your research ideas based on mentor feedback and research outcomes.

About

Inside Anthropic

Visit site →

Anthropic is a frontier AI research and product company, with teams working on alignment, policy, and security. We post specific opportunities at Anthropic that we think may be high impact. We do not necessarily recommend working at other positions at Anthropic. You can read concerns about doing harm by working at a frontier AI company in our career review on the topic.