SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Sarvam AI's Work Agents team is building the infrastructure backbone for developing, evaluating, and serving autonomous agents at scale. This DevOps Engineer role bridges platform engineering and operations, requiring hands-on expertise across Kubernetes, cloud infrastructure, CI/CD, and networking.
You will own Kubernetes platform operations across multi-cluster environments, managing upgrades, node lifecycle, RBAC, namespace hygiene, and policy enforcement. You'll design and maintain CI/CD pipelines that take code from commit to production, implementing rollout patterns like canary and blue-green deployments with built-in rollback capabilities.
Cloud infrastructure management is core to the role—provisioning and maintaining AWS resources (primary) and Azure (secondary) via infrastructure-as-code using Terraform or Crossplane. You'll handle VPCs, subnets, IAM, compute, storage, and managed services with reproducibility and cost awareness.
Networking responsibilities include managing CNI configuration, ingress and service routing, network policies, and secure cross-cluster connectivity via VPN tunnels, peering, and transit paths. You'll build automation and internal tooling—CLIs, scripts, operators—that make the harness self-service and reduce toil for engineers.
Observability and reliability are ongoing concerns: maintaining metrics, logging, and tracing pipelines; setting and tuning alerts; and participating in incident response and postmortems.
Required: 3+ years in DevOps, SRE, or platform engineering with hands-on ownership of systems you've run and improved. Deep Kubernetes fluency (day-to-day operations, scheduler understanding, operators/Helm/controllers), solid AWS proficiency (VPC, EC2, EKS, IAM, storage, networking), and practical networking depth (CNI, ingress, service routing, network policies, VPN/tunneling). CI/CD experience building and maintaining pipelines (GitLab CI, GitHub Actions, Argo CD, Flux), strong software engineering fundamentals (Python or Go preferred), and a product mindset toward internal users.
Bonus: on-premise infrastructure experience, air-gapped or restricted-network environments, multi-cluster/multi-tenant Kubernetes in production, or open-source contributions to Kubernetes or infrastructure tooling.