SlipstreamJobsFresh Startup & VC-Backed Jobs

Engineering Manager, Cloud Infrastructure

Replit - Foster City, CA, United States - In-office - posted 2026-09-24

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Replit is hiring an Engineering Manager to lead the Cloud Infrastructure team, responsible for the shared infrastructure-as-code (IaC), networking, storage, compute, and service mesh platforms that power Replit's product and platform teams. This is a hands-on, platform-building role with production accountability. You will own the cloud-platform roadmap, translating product, platform, reliability, and security needs into sequenced outcomes while balancing foundational investment, lifecycle work, and delivery commitments. You'll make infrastructure repeatable and self-service by building maintained IaC interfaces for services, cells, connectivity, identities, and shared resources, enabling internal teams to provision infrastructure without bespoke coordination. Operational ownership is central: you'll own platform availability, upgrades, isolation, recovery, and incident remediation, maintaining clear SLOs and sustainable on-call coverage. You'll stay technically engaged by reviewing designs and production changes, debugging failure modes, and using AI coding tools to prototype and automate infrastructure changes with rigorous verification. You'll build and grow a high-ownership engineering team by coaching engineers, developing technical leaders, setting clear expectations, managing performance, and hiring thoughtfully. The role requires delegation of meaningful ownership as the team scales. Replit is the agentic software creation platform enabling anyone to build applications using natural language, with millions of users worldwide. The infrastructure underneath must make it straightforward to launch services, isolate workloads, and run reliable systems at scale. REQUIREMENTS: - Demonstrated engineering management: led and developed engineers, made prioritization and performance decisions, hired thoughtfully, and delivered through a team - Software-oriented infrastructure depth: built and operated cloud platforms or distributed systems; can reason across infrastructure code, Kubernetes, networking, service identity, and stateful dependencies - Safe-change and production judgment: owned consequential migrations and incidents; can explain failure modes and rollback limits; knows when simplifying a system is better than adding another platform - Platform-product and engineering judgment: understands internal customers, creates interfaces other teams adopt, and makes clear tradeoffs among reliability, developer autonomy, engineering effort, and workload efficiency NICE TO HAVE: - Experience with multi-tenant, cellular, regional, or dedicated enterprise infrastructure - Familiarity with GCP/GKE, Terraform or similar IaC systems, Cloudflare, Envoy/Istio, SPIFFE/SPIRE, and managed data services - Experience with large-fleet rightsizing, infrastructure consolidation, or migrating CI compute without disrupting developer workflows - Track record using AI tools to increase engineering output while preserving production safeguards

Similar roles