SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Rubrik is hiring a Senior Software Engineer to join the Enterprise AI team, an internal AI enablement engine focused on making AI adoption seamless across the organization. This is a high-autonomy, hands-on role where you'll design, build, and operate the platform and backend systems powering Rubrik's infrastructure at scale.
You will own end-to-end delivery of major platform initiatives, from design through deployment and post-launch success. Core responsibilities include deep ownership of Kubernetes infrastructure—clusters, networking, operators, container lifecycle, and multi-tenant orchestration. You'll design, develop, and optimize distributed services and cloud-native infrastructure on AWS and/or GCP for scale, reliability, and performance.
Beyond individual contribution, you'll drive engineering excellence through code quality standards, design reviews, automation, and CI/CD best practices. You'll collaborate across Product, AI, and Security teams to align architecture with business objectives, mentor engineers through architecture decisions and trade-offs, and partner with leadership to align engineering strategy with product roadmap.
Required qualifications include 9+ years of software engineering with deep backend and infrastructure focus. You must have strong production-level programming skills in Python and/or Go, deep hands-on Kubernetes experience (building and operating clusters, not just deploying), and proven experience designing and operating distributed systems in production. Cloud-native fluency across AWS and/or GCP is essential, along with infrastructure-as-code (Terraform) and CI/CD pipeline experience. You should be familiar with applied AI tooling and patterns—agentic AI tools like Claude and LiteLLM, AI gateways, and agent frameworks—and able to build backend services integrating with them. Strong system design judgment and clear cross-functional communication are critical.
Preferred experience includes observability stacks (Prometheus, Grafana, Datadog, OpenTelemetry), multi-cloud or hybrid infrastructure, API/AI gateways and policy frameworks (ABAC, OPA), service mesh or platform-as-a-service design, and a track record of improving engineering productivity at scale.