SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
OpenAI's Product & Platform teams are seeking a Technical Program Manager to lead the operating system for ChatGPT capacity and model deployment. This role sits at the critical intersection of product demand, model deployment, inference, research, fleet management, and capacity planning.
You will own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations. Your core responsibilities include building durable intake, prioritization, and decision mechanisms that connect product demand and model requirements to available serving capacity. You'll partner with product, research, inference, fleet, and capacity teams to develop scenarios, surface tradeoffs, and drive timely allocation decisions.
Key ownership areas include model deployment readiness and rollout planning, including serving-capacity allocation, launch sequencing, validation, and operational handoffs. You'll establish clear readiness gates, risk reviews, rollback criteria, and escalation paths for model deployments. You'll drive launch coordination through deployment and post-launch learning, turning recurring gaps and manual work into scalable tooling and operating practices.
You'll define and operationalize metrics for forecast accuracy, capacity utilization and headroom, deployment velocity, reliability, latency, quality, and user impact. You'll create concise, decision-ready communications that make dependencies, risks, capacity constraints, and launch choices clear to technical and product leaders.
Ideal candidates have led complex technical programs in infrastructure, distributed systems, capacity planning, model serving, or large-scale deployment environments. You can reason credibly about demand, supply, headroom, reliability, latency, and quality tradeoffs, and translate them into executable plans. You've built operating mechanisms or tooling that replaced fragmented, manual workflows with scalable systems and clear ownership. You're effective in high-ambiguity, constrained environments where priorities change and decisions require explicit tradeoffs. You build alignment across research, engineering, product, finance, capacity planning, and operations without relying on direct authority. You use metrics to guide decisions, identify bottlenecks, and demonstrate measurable improvements in throughput, predictability, or reliability. You communicate with precision and move comfortably between technical detail, operational execution, and executive-level decisions.
The role is based in San Francisco with a hybrid work model of 3 days in the office per week. OpenAI offers relocation assistance to new employees.