SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
SentinelOne is seeking a Staff Site Reliability Engineer to join its Government SRE team. This role sits at the intersection of technical reliability and compliant, efficient deployments into government environments. You will own both the operational excellence of government-regulated cloud infrastructure and the coordination of releases that meet FedRAMP, DoD, and NIST standards.
Key responsibilities include driving continuous software delivery and incident management for production issues, with rapid recovery and root cause analysis. You will lead the design and optimization of observability strategies, collaborating with application engineering teams to enhance monitoring and reduce alert noise. You will define, implement, and monitor SLOs, SLIs, and SLAs in alignment with business objectives.
You will design and maintain software solutions addressing operational, compliance, and pipeline challenges, and own the full lifecycle of government environment releases, driving process improvements for efficiency and reliability. You will partner cross-functionally with engineering, product, security operations, compliance, and leadership to align priorities and resolve challenges. A critical aspect of this role is ensuring all infrastructure and deployments meet FedRAMP, government regulations, and industry standards while maintaining required documentation and risk assessments.
Ideal candidates bring 8+ years of SRE, DevOps, or Infrastructure Engineering experience for SaaS products, with 4+ years running operations at scale. You should have 2+ years of production experience with container orchestration (Kubernetes preferred) and continuous delivery. Strong understanding of compliance frameworks (FedRAMP, DoD, NIST 800-53, NIST 800-137) is essential. Multi-cloud experience with AWS/GCP (AWS preferred) is required. Proficiency in at least one programming language (Python, Go, Ruby) and bash scripting is expected. Familiarity with GitOps, Infrastructure as Code (Terraform or Pulumi), and deployment strategies (blue-green, rolling, canary) is important. Experience with observability stacks (Prometheus, Grafana, ELK, OpenTelemetry) and incident management processes is required. Proven background implementing FedRAMP, security, and compliance processes is a strong plus.
U.S. Citizenship and a work location in the United States are required due to Federal Government contract requirements. FedRAMP staff may be subject to customer or third-party background checks up to and including Secret Clearance if required.