SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Site Reliability Engineer

Sanity - Remote - Hybrid

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Sanity.io is a modern content operating system that replaces legacy CMS platforms, treating content as data so teams can maintain a single governed source of truth and adapt it across websites, apps, workflows, and AI agents. The company recently raised $85M in Series C funding and serves customers including SKIMS, Figma, Riot Games, Anthropic, and Nordstrom. As a Senior Site Reliability Engineer, you will design, build, and operate the shared platform foundations that enable Sanity's engineering teams to ship reliably at scale. The infrastructure handles approximately 75,000 requests per second across global infrastructure spanning multiple continents. Key responsibilities include: - Design and build GCP infrastructure, Kubernetes clusters, networking, routing, CI/CD pipelines, and observability systems - Diagnose and troubleshoot complex distributed systems running at high request volume - Ensure comprehensive observability and analyze stack behavior - Contribute to infrastructure modernization efforts including edge, caching, and gateway layers on Fastly - Raise reliability standards through improved dashboards, alert severity tuning, paging standards, on-call readiness, and incident response processes - Build golden paths and production readiness checks that make deployment safe and straightforward for engineers - Mentor engineers and raise technical standards through code review, design review, and pairing - Participate in on-call rotation and support developer on-call rollout You bring 5+ years of SRE/DevOps experience including on-call rotation responsibilities. You have hands-on experience with Kubernetes, cloud-based application scaling and high availability, CI/CD pipeline development, and observability stacks like Prometheus. You're comfortable working across CDNs, edge infrastructure, gateways, and caching layers. You approach infrastructure design analytically, diagnose issues systematically, and optimize for reliability. You communicate effectively during high-pressure incidents and build systems that improve production health over time. An open but thoughtful approach to new technologies is essential. Sanity is a 200+ person company with a global, multicultural team. The role requires US-based location with reasonable overlap to European engineering hours.

Similar roles