SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 131,000 - 164,000 / annual
Diligent is seeking a Staff Site Reliability Engineer to design, deploy, and operate VMware-based private cloud infrastructure powering mission-critical SaaS products globally. This is a hands-on technical leadership role where you'll own complex infrastructure spanning Linux, Windows Server, networking, storage, and automation frameworks across multiple global datacenters.
Key responsibilities include architecting and optimizing VMware vSphere environments; designing automation using PowerShell/PowerCLI, Ansible, and Python to reduce toil and increase reliability; administering and hardening Linux (RHEL/CentOS/Ubuntu) and Windows Server systems; managing Active Directory for hybrid on-premises and cloud authentication; partnering with network and security teams on firewalls, VPNs, storage, and load balancers (F5 BIG-IP, AVI/NSX); and mentoring engineers while setting technical direction for long-term reliability and automation strategy.
Required qualifications: 10+ years in systems or infrastructure engineering with large-scale enterprise or SaaS datacenter experience; deep hands-on expertise with VMware vSphere in production; strong Linux administration skills including performance tuning and hardening; solid Windows Server and Active Directory experience; proven track record building automation with PowerShell/PowerCLI, Ansible, or Python; understanding of storage (SAN/NAS), TCP/IP networking, DNS, VPNs, firewalls, and monitoring; collaborative problem-solving mindset with ability to lead complex incidents and operate in high-availability on-call environments.
Desirable: experience with enterprise storage (Pure Storage) or compute (Cisco UCS); familiarity with Terraform, Jenkins, Azure DevOps, or infrastructure-as-code tools; exposure to security hardening and compliance frameworks (CIS, NIST, ISO 27001).