SlipstreamJobsFresh Startup & VC-Backed Jobs

Director, Site Reliability Engineering

Okta - Bengaluru, India - Hybrid - posted 2026-08-25

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Okta is seeking a Director of Site Reliability Engineering to lead and build a high-caliber SRE organization based in Bengaluru, India. This role oversees the SRE teams supporting Okta's production platform, which authenticates, authorizes, and provisions millions of users daily across AWS infrastructure with 99.999% availability targets. Key responsibilities include building and leading the India-based SRE organization, defining and executing India SRE strategy aligned with global reliability goals, and partnering with global engineering, product, and infrastructure leaders to deliver resilient, scalable, and secure services. You will lead post-incident reviews, drive root-cause analysis, and ensure long-term corrective actions while participating in incident management and on-call rotations. The role emphasizes implementing automation and observability to reduce manual toil and improve operational efficiency. You will drive adoption of modern infrastructure practices including infrastructure as code (Terraform), container orchestration (Kubernetes), and AI within the infrastructure organization. Hiring, mentoring, and developing top SRE talent across India is critical, with a focus on building a strong engineering culture centered on reliability and innovation. You will foster collaboration across time zones with U.S. and EMEA teams, promote continuous learning and knowledge sharing, and maintain deep knowledge of industry best practices and evolving technologies. The role includes managing service and business expectations, prioritizing resource allocation, and accelerating velocity through robust platforms, powerful tooling, and self-service capabilities. You will also improve SDLC processes for cloud infrastructure as code, including CI/CD pipeline maturity and change/release management. Required qualifications include 16+ years in site reliability, infrastructure, or production engineering; 8+ years in technical leadership and people management including managing managers; 4+ years running SRE organizations supporting SaaS/Cloud services on public cloud (preferably AWS); and strong expertise in cloud-native architectures, Kubernetes, Terraform, and CI/CD pipelines. A Computer Science degree or equivalent experience is required.

Similar roles