SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
OpenAI's Safety Systems team is hiring a Model Policy Manager to shape model behavior for U.S. government use, with a focus on national security applications. You will define nuanced policies and translate them into training and evaluation criteria, helping frontier models navigate high-stakes scenarios while preserving their usefulness and capabilities.
The role is based in OpenAI's San Francisco office with a hybrid model: three days in the office per week with optional work from home on Thursdays and Fridays. Relocation support is available for new employees.
Key Responsibilities:
- Develop model policies that guide safe and useful behavior in national security contexts
- Build evaluations, identify policy gaps and model failures, and use findings to improve policies and training
- Work with research, engineering, and domain experts to support safe, reliable deployment of frontier models
- Translate complex safety questions into clear, practical policies
- Work hands-on with model data and evaluations
About the Team:
Safety Systems manages the complete lifecycle of safety efforts for OpenAI's frontier models, ensuring responsible deployment and positive societal impact. The work spans system-level safeguards, model training, evaluation, and red-teaming to mitigate misuse and misalignment. The Model Policy team specifically designs policies that define safe model behavior in real-world environments.
Requirements:
- Relevant experience in AI safety, policy, or risk assessment
- Strong judgment and ability to turn complex safety questions into clear, practical policies
- Technical fluency to work hands-on with model data and evaluations
- Motivation aligned with OpenAI's mission and responsible AI use in safety-critical settings
- Active TS/SCI clearance or equivalent (preferred but not required)