SlipstreamJobsFresh Startup & VC-Backed Jobs

Site Reliability Engineer

Anduril - Waltham, MA, United States - In-office - posted 2026-07-27

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Anduril Industries is seeking a Site Reliability Engineer to join the Imaging team, which builds and deploys state-of-the-art camera and sensor systems for U.S. and allied military applications. This is a field-focused SRE role, distinct from traditional cloud SRE work. You will be the frontline operator responsible for keeping deployed imaging systems operational in the field. You'll serve as the escalation point for field personnel and customer-support teams when deployed systems malfunction. The work spans the full stack—from bare-metal hardware and firmware to networked services and cloud integrations—and you own systems through deployment and into the field. Key responsibilities include: - Owning the health and uptime of deployed imaging systems, triaging and diagnosing issues that could originate anywhere in the stack (network, calibration, upgrades, sensor hardware). - Running point on escalations from support channels and customer-support pipelines, serving as the deep-expertise backstop. - Converting recurring field issues into durable runbooks, diagnostics, and self-service tooling to reduce repeat problems. - Maintaining the boundary with product engineering: you own everything except code fixes; when you identify a genuine software defect, you reproduce and document it cleanly for handoff. - Feeding field reliability insights back into product design to improve observability, upgrade safety, and failure gracefully. - Approximately 15% travel for field support and deployment windows. You'll be the second SRE on a small, high-trust team, working directly with the lead SRE, with real opportunity to shape how imaging reliability and support scale. Required qualifications: 3+ years in SRE, DevOps, field/systems engineering, or production support of deployed hardware/software systems with real post-ship ownership; strong Linux fundamentals and networking troubleshooting (IP, routing, VPNs, field environments); ability to diagnose and resolve cross-boundary issues without full visibility; comfort with structured on-call rotation including after-hours/weekend coverage; strong written and verbal communication for remote troubleshooting and documentation; eligibility to obtain and maintain a U.S. Secret clearance.

Similar roles