SlipstreamJobsFresh Startup & VC-Backed Jobs

Software Engineer, Reliability (SRE)

Veeam Software - Pune, Maharashtra, India - In-office - posted 2026-08-12

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Veeam is seeking a Senior Software Engineer, Reliability to join its SRE team in Pune. In this hands-on technical leadership role, you will guide senior engineers, influence product development teams, and ensure systems are built to be reliable, scalable, and observable from inception. You will drive strategic SRE initiatives across Veeam's global platform, mentor engineers in SRE practices, and define architectural best practices. Key responsibilities include designing and evolving highly available, fault-tolerant infrastructure on public clouds (primarily Azure); establishing and maintaining SLIs, SLOs, and error budgets; leading incident response and blameless postmortems; and driving adoption of deep observability practices with comprehensive telemetry, logs, metrics, and tracing. You will develop automation and self-healing tools to reduce operational toil, contribute to infrastructure-as-code, CI/CD systems, and deployment automation. You'll integrate chaos engineering tools to validate reliability assumptions, implement testing strategies and canary deployments, and participate in on-call rotations. Collaboration is central: you'll embed within product and platform teams to champion reliability from design through delivery, mentor engineers globally, and advocate for DevOps/SRE best practices. Required: 5+ years hands-on software engineering experience with at least 2 years in Site Reliability, Platform Engineering, or similar roles. Deep experience with public cloud providers (Azure preferred), strong programming skills (JS, Node, TypeScript, Go, Java, C#), and proven track record delivering monitoring and observability tooling (Prometheus, Grafana, OpenTelemetry). Experience with IaC tools (Terraform/Pulumi), container orchestration (Kubernetes), distributed systems, cloud networking, and cloud-native design required. Excellent communication and collaboration skills across geographies essential. Bonus: Large-scale B2B SaaS platform experience, chaos engineering and resilience testing background, compliance framework familiarity (ISO, SOC 2, GDPR, FEDRAMP/CMMC).

Similar roles