HireHireAnthropic jobs › Safeguards Enforcement Lead, Cyber Harms

Safeguards Enforcement Lead, Cyber Harms

Apply for this role or explore on the map →

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As an Enforcement Lead, you will be responsible for managing and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your work will center on developing strategic enforcement frameworks for flagged activity related to cyberattacks, malware development, and offensive exploitation. Additionally, you will manage a team of Cyber Enforcement Analysts and contractors implementing this enforcement strategy. 

Safety is core to our mission, and you'll help uphold policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way.

Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, technical, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays.

Key responsibilities

Minimum qualifications

Preferred qualifications