Anthropic is actively hiring for various roles in AI safety and security to address potential catastrophic outcomes associated with advanced AI technology. This recruitment effort aligns with the company’s ongoing updates to its Frontier Safety Roadmap and Responsible Scaling Policy, aimed at developing AI responsibly. Additionally, Anthropic has expanded its Fellows Program to focus on AI safety, with new cohorts starting in mid-2026, highlighting its commitment to reducing catastrophic risks in the field.

Anthropic: Anthropic is an AI safety and research company focused on developing reliable, interpretable, and steerable artificial intelligence systems. The organization prioritizes addressing potential risks from advanced AI through dedicated research and policy efforts. It is currently expanding recruitment for specialized safety, security, and fellowship roles explicitly aimed at reducing catastrophic risks from AI.

Safety Roadmap: Anthropic continues to update its Frontier Safety Roadmap and Responsible Scaling Policy to provide guidance on developing advanced AI responsibly.
Fellows Program: Anthropic has expanded its Fellows Program to include dedicated workstreams in AI safety and security, with cohorts starting in mid-2026 and applications emphasizing motivation to reduce catastrophic risks.
AI Safety Hiring: Anthropic is actively recruiting across AI safety, security, and related policy roles as part of efforts to mitigate potential catastrophic outcomes from advanced AI.