SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Scale AI is seeking an Engineering Manager for Infrastructure to lead a team responsible for the full deployment lifecycle of AI-powered applications in live customer environments. The role bridges AI research and production, turning innovative prototypes into scalable, high-performance enterprise solutions.
You will own the infrastructure roadmap aligned with business and engineering priorities, leading the design and implementation of scalable, secure, and reliable infrastructure systems. Key responsibilities include setting and maintaining SLAs/SLOs for platform uptime and performance, managing the engineering team to drive technical delivery, and designing backend services for AI-driven applications including AI agents and evaluation tooling.
The role emphasizes staying close to customer experience, translating real-world deployment challenges into platform improvements and scalable playbooks. You'll work cross-functionally with customers, GTM, product, and infrastructure teams to ensure every deployment is production-ready and secure.
You will inspire and mentor engineers, influence team culture and values, and collaborate with product, security, and engineering leadership on strategic alignment. The team works on interactive AI applications, enterprise SaaS products, and platform capabilities that help businesses leverage AI effectively.
Required qualifications include 5+ years of relevant experience with 2+ years managing infrastructure or platform teams. You must have proven experience with cloud platforms (AWS, GCP, Azure, OCI), Kubernetes deployments on on-prem infrastructure, CI/CD pipelines, infrastructure-as-code tools like Terraform, and production environment management. Experience with monitoring, alerting, incident response, modern developer platforms, and network engineering is essential.