SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
The Search Product Infrastructure team at OpenAI builds the systems powering search experiences across ChatGPT. This team partners with model development, inference infrastructure, specialized search, and indexing teams to bring advances in models and retrieval into production at scale.
As a Senior Engineer on this team, you will design, build, and operate systems that connect models with search at ChatGPT scale. You will own projects from technical design through launch and iteration, working closely with researchers and partner engineering teams to evolve the architecture as models, product capabilities, and demand grow.
Key responsibilities include:
- Design and evolve services that coordinate search classification, retrieval, ranking, and model inference, bringing new capabilities into production.
- Improve end-to-end latency, throughput, and infrastructure efficiency through profiling, caching, request routing, and capacity planning, making informed tradeoffs between search quality, reliability, and compute cost.
- Build experimentation tooling and automation to run reproducible A/B tests, define success metrics, and measure product impact. Use shadow traffic and load testing to validate system behavior and support safe production rollouts.
- Own production reliability through observability, resilient fallback behavior, incident response, and automation that improves launch safety and reduces operational toil.
- Build search APIs and tool interfaces that enable models and agents to retrieve information reliably while respecting access controls and preserving source attribution.
The role is based in San Francisco with a hybrid work model of 3 days in the office per week. Relocation assistance is offered to new employees.
Qualifications:
- Significant experience designing, building, and operating large-scale distributed systems, with depth in performance, reliability, or resource efficiency.
- Experience in one or more of: search, information retrieval, ML infrastructure, inference serving, or other high-throughput online systems.
- Strong programming and systems debugging skills; comfort working across languages and unfamiliar parts of a production stack.
- Ability to turn ambiguous product or research needs into clear technical plans, align partners across teams, and carry projects through deployment and measurable results.
- Comfort learning across systems and ML, investigating unfamiliar problems, and sharing technical decisions clearly with others.