SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Anthropic is seeking a Safeguards Enforcement Lead to manage the User Well-Being team's enforcement operations. This is a management role responsible for overseeing child safety, mental health, abuse and exploitation, and age assurance enforcement workflows.
You will manage a team of individual contributors conducting content review across multiple policy areas. Key responsibilities include serving as the primary point of contact for review partners, designing and improving enforcement workflows to scale effectively while maintaining accuracy and consistency, and partnering with Engineering and Data Science teams to optimize detection models and automated enforcement systems.
You will develop and maintain internal documentation, decision trees, and review guidelines that enable accurate enforcement at scale. The role requires staying current with emerging AI policy enforcement best practices, evolving legal frameworks, and technological developments to inform workflow improvements.
Additionally, you will identify and report trends in misuse patterns to internal stakeholders including Policy, Legal, and Trust & Safety leadership, and coordinate reporting obligations to external bodies such as NCMEC in accordance with applicable law and company policy.
Minimum qualifications include experience managing teams in the User Well-Being space, direct exposure to child safety, mental health, abuse and exploitation, and age assurance harm areas through trust & safety or content moderation roles. You must have experience managing or coordinating content review operations, scaling policy enforcement workflows, proficiency in SQL or data analysis tools, and the ability to identify emerging risks and communicate findings to cross-functional teams.
Preferred qualifications include deep expertise in child safety and CSEA, experience with NCMEC or equivalent reporting bodies, familiarity with relevant legal frameworks (CSAM reporting, KOSA, COPPA), experience with generative AI products, knowledge of trauma-informed support structures for content reviewers, Python proficiency, and familiarity with hash-matching or perceptual hashing technologies.
Important context: This role involves regular exposure to explicit sexual content involving minors and potentially violent or psychologically disturbing material. Anthropic provides wellness resources and support for team members in this area.