SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 146,000 - 194,000 / annual
Anduril Industries is a defense technology company transforming U.S. and allied military capabilities through advanced technology. The company's systems are powered by Lattice OS, an AI-powered operating system that integrates thousands of data streams into a real-time, 3D command and control center.
The CorpTech Platform team is the internal engineering force multiplier behind Anduril's corporate systems, driving strategic investment in data platforms, software platforms, QA and release excellence, ERP engineering, and AI infrastructure. The team builds foundations powering CorpOS (enabling Finance and Growth operations) and ArsenalOS (the digital backbone of Anduril's hardware enterprise). The organization is also driving Anduril toward becoming an autonomous enterprise through initiatives like the Autonomous Software Factory, rethinking how software is built, tested, deployed, and evolved by integrating AI directly into the engineering lifecycle.
In this role, you will act as the quality and reliability lead within a small, senior AI development pod focused on solving novel, high-impact business and engineering problems. You will design and implement test strategies, validation approaches, and release readiness criteria for AI-enabled software, automation, and agentic workflows. You will partner closely with software and AI engineers to identify failure modes early across code, prompts, models, integrations, infrastructure, and user workflows. You will build confidence in fast-moving solutions through automated testing, observability, instrumentation, environment design, and operational safeguards.
Key responsibilities include helping define practical standards for reliability, resiliency, debugging, and production operations in an AI-first development model; contributing directly in code, infrastructure, and tooling to improve delivery confidence, developer feedback loops, and production stability; supporting launch readiness, production issue response, root cause analysis, and continuous improvement for front-facing systems; and evangelizing strong engineering discipline in quality assurance, release engineering, infrastructure hygiene, and incident prevention without sacrificing speed.
REQUIREMENTS:
- 5+ years of experience in software engineering, site reliability engineering, quality engineering, infrastructure engineering, or a closely related role in a fast-paced environment
- Demonstrated experience building or operating reliable production software systems, including ownership of testing, observability, deployment confidence, and operational readiness
- Strong technical fluency in modern software architectures, APIs, distributed systems, CI/CD, and cloud or platform infrastructure
- Experience working with frontier AI tooling, AI coding assistants, LLM-enabled applications, or agentic systems, including awareness of unique quality and reliability challenges
- Ability to move between hands-on implementation and systems-level quality strategy, with sound judgment on where rigor is required versus where speed is appropriate
- Strong debugging, root cause analysis, and incident response instincts, with a bias toward preventing classes of failures
- Excellent written and verbal communication skills, with ability to influence senior engineers and cross-functional stakeholders on quality and reliability trade-offs
- Degree in Computer Science, Information Systems, Engineering, or related technical field, or equivalent practical experience
- U.S. Person status required (position needs to access export controlled data)
PREFERRED QUALIFICATIONS:
- Experience supporting AI-first or AI-accelerated software development teams, including designing quality controls around non-deterministic system behavior
- Experience with automated testing strategies spanning unit, integration, end-to-end, performance, and reliability testing
- Experience with infrastructure as code, cloud platforms, observability stacks, release engineering, and production operations
- Experience in hyper-growth startup-like environments, balancing speed, ambiguity, and engineering rigor
- Familiarity with enterprise systems and business process domains such as ERP, MES, WMS, CRM, finance systems, or manufacturing systems
- Eligible to obtain and maintain a U.S. Secret security clearance