SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
LeoLabs is building a living map of activity in space through a proprietary global radar network and AI-enabled analytics platform. The company collects millions of measurements daily on 25,000+ objects in low Earth orbit, providing space domain awareness and satellite operations intelligence to commercial and government missions.
As a Senior Site Reliability Engineer, you will bridge development and operations to ensure LeoLabs' mission-critical systems are scalable, reliable, and efficient. You'll design and maintain scalable infrastructure, set up comprehensive monitoring and incident response systems, and develop automation tools for deployment and system health checks. Key responsibilities include capacity planning to forecast future needs, collaborating with development teams to enhance product reliability, creating system documentation, participating in on-call rotations for 24/7 support, and implementing security best practices across all systems.
You'll work with a modern engineering stack including Python/Go, AWS/Azure cloud services, Docker/Kubernetes containerization, Terraform infrastructure-as-code, CI/CD tools like GitHub Actions, and monitoring platforms such as Grafana and Datadog. The role requires 5+ years of SRE, DevOps, or related experience, with strong expertise in distributed systems, microservices architecture, and database technologies.
Success milestones include completing onboarding within one month, independently deploying infrastructure changes within three months, optimizing deployment pipelines and cloud costs within six months, and leading cross-functional reliability initiatives and mentoring junior engineers within twelve months. The company values problem-solving, collaboration, and the ability to obtain U.S. personnel security clearance (active TS/SCI clearance preferred).