SlipstreamJobsFresh Startup & VC-Backed Jobs

Software Engineer, Fleet Infrastructure

Anduril - Boston, MA, United States - Hybrid - posted 2026-07-27

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Salary: USD 166,000 - 220,000 / annual

Anduril Industries is seeking a Software Engineer to join the Observability team, which builds and operates the company's production telemetry systems for ground and cloud nodes. The role focuses on designing and implementing core developer SDKs for metrics aggregation, structured logging, and tracing; building real-time data pipelines connecting edge and cloud environments; and operating centralized monitoring and alerting infrastructure. Key responsibilities include: • Build and operate a robust, high-availability observability plane serving 100s of environments with strict uptime requirements and graceful scaling • Design developer-facing SDKs and tooling for instrumenting mission-critical services across heterogeneous fleets (Go, Rust, C++, Python) • Partner closely with internal engineering teams to understand needs, identify system gaps, and proactively improve observability offerings • Enable metrics-driven engineering and seamless production incident investigation, including support for agentic workflows and cross-service root-cause analysis • Support both cloud-connected and offline/airgapped environments with seamless user experience • Collaborate with Robotics Data Foundation and Fleet Management teams on vehicle telemetry and software delivery infrastructure The Observability team operates mission-critical systems where reliability is paramount—Anduril's products serve high-stakes military environments where failures have life-or-death consequences. You'll work on infrastructure that aggregates signals from edge to cloud across Anduril's core ground systems and Kubernetes infrastructure. Required: 3+ years software engineering experience, bachelor's degree in Computer Science or related field (or equivalent), familiarity with container orchestration (Docker, Kubernetes) and cloud platforms (AWS, GCP, Azure), experience building systems in Go, C++, Rust, or Java, and U.S. Person status for access to export-controlled data. Preferred: Experience with observability stack components (ClickHouse, Victoria Metrics, Prometheus, Grafana, ELK), on-call rotation support for high-availability systems, FedRAMP compliance knowledge, and deployment experience in government cloud enclaves (AWS GovCloud, Azure Government).

Similar roles