SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Site Reliability Engineer -

Okta - Bellevue, WA, United States - Hybrid - posted 2026-02-12

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Okta is seeking a Senior Site Reliability Engineer to join its Security and Data Systems team in Bellevue, Washington. This hybrid role blends software engineering with systems administration, focusing on building and maintaining highly reliable, scalable, and secure infrastructure for a large-scale SaaS platform. Key responsibilities include designing and maintaining core infrastructure for security SaaS offerings, ensuring high availability, performance, and scalability. You will develop robust automation using code to eliminate toil and ensure consistency across environments, from infrastructure provisioning to application deployment and incident response. The role emphasizes a security-first mindset, requiring close collaboration with security teams to embed compliance and security best practices into all processes and infrastructure. You will participate in on-call rotations as a primary responder for critical incidents, leading root cause analysis and implementing preventative measures. Collaboration is central to the role—you'll partner with development, data science, and security teams to provide expert guidance on architectural decisions and best practices. Required qualifications include strong production-level coding skills, deep experience with Terraform for infrastructure as code, familiarity with modern CI/CD practices (particularly Spinnaker), and expertise in containerization and orchestration with Kubernetes. Direct experience with large-scale data systems, specifically Snowflake, is essential. Additional valuable skills include experience with database schema management tools like Flyway, excellent analytical and problem-solving abilities, and a proactive approach to identifying potential issues. Experience or strong interest in AI/ML applications for reliability, security, and operational efficiency (AIOps, predictive analysis) is a plus. The role requires in-person onboarding and travel to the Toronto, Canada office during the first week of employment. This is a career-defining opportunity to work on complex challenges with real-world stakes at a company focused on securing AI and identity infrastructure.

Similar roles