SlipstreamJobsFresh Startup & VC-Backed Jobs

Site Reliability Engineer

Anduril - Waltham, MA, United States - In-office - posted 2026-09-01

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Anduril Industries is seeking a Site Reliability Engineer to join the Imaging team, which builds and deploys state-of-the-art camera and sensor systems for U.S. and allied military applications. This role is the frontline for keeping fielded imaging systems operational in real-world deployed environments. You will own the reliability and uptime of deployed imaging systems, serving as the escalation point for field personnel and customer-support teams when systems malfunction. The work spans the full stack—from networking and firmware to cloud integrations—and requires diagnosing issues that could originate anywhere in the system architecture. You'll triage live incidents, walk field operators through troubleshooting procedures, and systematically convert recurring problems into durable runbooks and self-service diagnostics. Key responsibilities include: managing fielded system health and incident response; serving as the deep-expertise backstop for support escalations; building runbooks and diagnostics to reduce repeat issues; maintaining the boundary between field support and product engineering by cleanly reproducing and documenting software defects; and feeding field reliability insights back into product design for better observability and graceful failure modes. You will be the second SRE on a small, high-trust team, working directly with the lead SRE with significant autonomy to shape how imaging reliability scales. Approximately 15% travel is expected for field support and deployment windows. Required qualifications: 3+ years in SRE, DevOps, field/systems engineering, or production support of deployed hardware/software systems with real post-deployment ownership; strong Linux fundamentals and networking troubleshooting skills (IP, routing, VPNs, constrained environments); demonstrated ability to diagnose cross-boundary issues (networking, services, hardware) with incomplete visibility; comfort with structured on-call rotations including after-hours/weekend coverage; strong written and verbal communication for remote troubleshooting and documentation; eligibility to obtain and maintain a U.S. Secret clearance.

Similar roles