SlipstreamJobsFresh Startup & VC-Backed Jobs

Safeguards Enforcement Lead, Cyber Harms

Anthropic - Washington, DC, United States - Hybrid - posted 2026-08-29

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Salary: USD 285,000 - 330,000 / annual

Anthropic is seeking a Safeguards Enforcement Lead to manage and execute enforcement actions across its AI products, with a focus on detecting and mitigating misuse of AI systems for malicious cyber operations. You will lead a team of Cyber Enforcement Analysts and contractors, developing strategic enforcement frameworks to address cyberattacks, malware development, and offensive exploitation attempts. Key responsibilities include managing your enforcement team and overseeing the vision of cyber enforcement strategy; creating detection and mitigation strategies for AI system misuse in cyber operations; collaborating with stakeholders on novel, ambiguous, or high-severity cases; working with the Safeguards Policy Design Team to identify and address policy gaps; partnering with Engineering and Data Science teams to build tooling and measurement systems; and staying current with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape. You must have experience as a people manager and deep cybersecurity expertise, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research. Content review, abuse investigations, or policy enforcement at scale is required. Proficiency in SQL and/or Python for data analysis and threat detection is essential. You should be able to identify emerging risks and communicate findings to diverse stakeholders (Product, Policy, Engineering, Legal). Experience with generative AI products and writing effective prompts for content review is required. Preferred qualifications include trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence experience at a technology or AI company; familiarity with large language models and how AI could be misused for cyber operations; experience with abuse monitoring programs or enforcement review systems; understanding of product policy implementation at scale in content moderation; and experience working with government agencies or regulated environments. Note: This role involves exposure to explicit content spanning violent, technical, and psychologically disturbing topics. You may need to respond to escalations during weekends and holidays. The role is hybrid with at least 25% office time expected, with travel required.

Similar roles