SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 150,000 - 250,000 / annual
Foundry Robotics is building an AI-native robotics manufacturing company focused on deploying advanced assembly and production capability for leading robotics companies and national-security-critical hardware. You will work on the infrastructure that services, factory-floor software, and ML training run on: AWS, Kubernetes, the network boundary between them, the deployment pipeline, GPU workload scheduling, and monitoring systems.
This is a unique position with significant scope, spanning traditional containerized software platforms, robotics, networking, and AI infrastructure. Unlike internet companies, you will work on capabilities that support real hardware. Unlike traditional robotics, this platform orchestrates the manufacturing process and continued operation of robots themselves.
You will work on a small team with broad surface area. Expect to move between infrastructure as code, cluster operations, GPU scheduling, and edge-device support in the same week. The role requires hands-on work across cloud backends, on-premises services, and edge devices deployed on the factory floor. You will ensure that operators, supervisors, and managers have fast, reliable, and intuitive interfaces—even when network conditions aren't perfect and the environment is demanding.
The company is committed to being deeply embedded in the U.S. industrial base, building adaptive robotic assembly systems that make American manufacturing scalable, resilient, and competitive.
REQUIREMENTS:
- 3+ years running production infrastructure with hands-on Kubernetes
- Fluent in Terraform and GitOps; experience with infrastructure CI/CD where merge means apply, with knowledge of safe deployment practices
- Solid networking fundamentals: VPCs, routing, VPN, DNS, TLS, overlay networks, and ability to debug routing problems with packet captures
- Comfortable in Go or Python for tooling and glue; able to read production service code well enough to debug it
- Experience operating an observability stack with a track record of making alerts trustworthy
- Security-minded by default: least-privilege access, no public endpoints, secrets from config, and reflexes to catch committed credentials
STRONG PLUS:
- Experience running Kubernetes on-premises
- Edge or fleet experience: arm64 Linux devices, over-the-air rollout, offline-tolerant design
- WireGuard-based overlay networks at organization scale
- Time spent near robots or industrial hardware
- Experience in an ITAR or CMMC environment