SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Veeam, the Data and AI Trust Company, is seeking a Senior Software Engineer, Reliability (SRE) to join its Site Reliability Engineering team in Pune. This is a hands-on technical leadership role where you will guide senior engineers, influence product development, and ensure systems are built for reliability, scalability, and observability at scale.
In this role, you will design and evolve infrastructure to be highly available, fault-tolerant, and scalable across public clouds (primarily Azure, with expansion plans). You will establish and maintain SLIs, SLOs, and error budgets that define reliability objectives, lead incident response and blameless postmortems, and drive organizational learning from incidents.
You will drive adoption of deep observability practices, ensuring comprehensive telemetry, logs, metrics, and tracing. You'll develop automation and self-healing tools to reduce operational toil and support fleet management strategies. You'll contribute to infrastructure as code, CI/CD systems, deployment automation, and scalable configuration management. You will integrate monitoring and chaos engineering tools to validate reliability assumptions under load and failure conditions, and implement testing strategies and canary deployments to protect production.
Collaboration is central to this role. You will embed within product and platform teams to champion reliability from design through delivery, mentor engineers globally, and advocate for DevOps/SRE best practices across the organization.
Required qualifications include 5+ years of hands-on software engineering experience with at least 2 years in Site Reliability, Platform Engineering, or similar roles. You need deep experience building systems on public cloud providers (Azure preferred), strong programming skills in JavaScript, Node, TypeScript, Go, Java, C#, or similar languages, and proven track record delivering monitoring, alerting, and observability tooling such as Prometheus and Grafana. Experience with infrastructure-as-code tools like Terraform/Pulumi, container orchestration (Kubernetes), distributed systems, cloud networking, and cloud-native design is essential.
Bonus qualifications include experience on large-scale B2B SaaS platforms, chaos engineering, resilience testing, performance testing, and familiarity with compliance frameworks (ISO, SOC 2, GDPR, FEDRAMP/CMMC).
Veeam offers 18 paid vacation days plus 4 global VeeaMe Days for self-care, 24 paid volunteer hours annually, private medical coverage for you and up to four dependents, life/accident/disability insurance, annual flexible wellbeing allowance, free confidential counseling via EAP, meal/fuel/transportation benefits, daycare reimbursement, and learning opportunities through LinkedIn Learning, O'Reilly, mentoring, and workshops.