SlipstreamJobsFresh Startup & VC-Backed Jobs

Strategic Projects Lead

Patronus AI - San Francisco, CA, USA - In-office - posted 2026-08-24

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Patronus AI is a frontier AI lab developing simulation research and infrastructure to accelerate progress toward human-aligned AGI. The company is behind influential research in AI evaluation including FinanceBench, Lynx, SimpleSafetyTests, CopyrightCatcher, and Humanity's Last Exam. Founded by former researchers from Meta AI, Amazon AGI, and Google, Patronus serves foundation model labs and Fortune 500 enterprises like Adobe, backed by top-tier investors including Lightspeed Venture Partners and Notable Capital. As Strategic Projects Lead, you will lead the delivery of high-quality simulations that define how AI systems are trained, evaluated, and improved. This is a highly autonomous role managing a team to build simulations of impactful real-world workflows, owning project execution, quality standards, and customer alignment from requirements through delivery. Key responsibilities include: leading end-to-end delivery of simulations and environments for real-world workflows; maintaining clear visibility into open tasks, delivery gaps, quality risks, and timeline confidence; serving as primary delivery interface with customers, aligning on requirements and managing expectations; partnering with technical, research, and QA leads to prioritize build efforts; defining and upholding objective quality standards for simulations, tasks, and environments; analyzing model behavior and failure modes to inform reward design and QA processes; and translating learnings into scalable systems and processes. You will work at the intersection of reinforcement learning, scalable oversight, and real-world workflow simulation. Your work directly influences how frontier models are developed, stress-tested, and deployed, informing how frontier labs design and train the next generation of agents for long-horizon tasks. Required qualifications: BS, MS, or equivalent in Computer Science, Machine Learning, Engineering, Mathematics, or related technical field; >1 year experience in human data space working on RL environments and task design at a top-tier company; strong technical foundation with ability to learn quickly and drive execution.

Similar roles