SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Fireworks is a Series D AI infrastructure platform enabling companies to build, train, and serve specialized AI models tailored to their data and workflows. Backed by top-tier investors including NVIDIA, Sequoia, Benchmark, and AMD, the company is valued at $17.5B and founded by PyTorch veterans.
As a Member of Technical Staff in Cloud Infrastructure, you will design and develop core backend systems powering Fireworks' high-performance generative AI platform. You'll focus on ensuring efficiency, scalability, and stability across AI workloads, working on challenges ranging from low-latency inference to scalable model serving.
Key responsibilities include:
- Design and build core backend software components for efficiency, scalability, and stability
- Conduct design and code reviews, collaborating across engineering and product teams
- Continuously analyze and optimize infrastructure efficiency for AI workloads (compute, storage, networking)
- Tackle hard problems at the forefront of AI infrastructure
You bring 3+ years of experience in ML infrastructure (PyTorch, Vertex AI, SageMaker, or similar), with proven expertise building, scaling, and optimizing enterprise-grade machine learning systems. A Bachelor's degree in Computer Science, Computer Engineering, or equivalent practical experience is required. Advanced degrees (Master's/PhD) are preferred.
This is an opportunity to work with world-class engineers and AI researchers on bleeding-edge technology that impacts how businesses globally harness AI. The role emphasizes ownership, direct impact, and learning from industry veterans in a fast-growing, collaborative environment.