SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
OpenAI's Critical Harm Operations team is seeking a senior cybersecurity practitioner and operations strategist to lead the Cyber vertical within User Safety & Risk Operations. This senior individual contributor role focuses on building enforcement systems for frontier risk and material harm that are accurate, fast, defensible, and scalable.
You will drive the cyber operations operating model across domain priorities, standard operating procedures, escalation paths, quality health, vendor capability, and roadmap inputs. You'll serve as the senior cyber expert for complex or high-risk decisions across ChatGPT, API, Codex, agents, and emerging product surfaces. The role requires translating policy ambiguity, quality misses, appeals, and reviewer disagreement into clear decision rules, calibration examples, training, and tooling requirements.
Key responsibilities include building durable operating systems and quality loops (golden sets, holdouts, double-labeling, adjudication, error taxonomies, reviewer calibration, and automation evaluations), raising FTE and BPO capability through onboarding and certification, and diagnosing root causes using quality, appeals, SLA, backlog, and disagreement signals. You'll build hands-on solutions using SQL, Python, dashboards, LLM eval workflows, and lightweight automations to improve decision quality and reduce manual effort.
Success is measured by durable improvement in the operating model and the reviewers who run it, not primarily by cases closed. You'll partner with Policy, Integrity, Safety Systems, Security, Legal, Product, Engineering, and Investigations to operationalize changes and drive launch readiness.
Required qualifications include 8+ years of hands-on cybersecurity experience in offensive security, threat intelligence, incident response, security research, red teaming, application security, DFIR, malware analysis, or related fields. You must reason deeply about attacker tradecraft, vulnerability exploitation, credential abuse, malware, persistence, evasion, exfiltration, cloud or identity abuse, and ambiguous dual-use activity. Experience building or improving high-stakes operations, reviewer programs, QA systems, escalation workflows, or vendor/BPO programs is essential. Technical proficiency with SQL, Python, C/C++, JavaScript, PowerShell, Bash, APIs, and LLM tooling is required. Experience with trust and safety, platform abuse, cyber misuse of AI systems, or LLM safety is highly valued.