SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 220,000 - 255,000 / annual
Legion is a workforce management platform serving hourly workers and their employers. The SRE team is responsible for building and maintaining the tools, infrastructure, and services that power Legion's secure, highly scalable, and cost-effective AWS/Kubernetes-based cloud platform.
In this Principal Software Engineer, DevOps role, you will:
- Support Legion's public cloud platform utilizing multiple AWS services and containerization
- Develop infrastructure automation leveraging Terraform, Chef, Jenkins, and Golang
- Create and manage production alerts, respond to incidents, and conduct root-cause investigations
- Develop automated operational runbooks
- Support deployment of Legion's WFM solution during and outside regular office hours
- Participate in on-call rotation
- Build and operate internal AI agents that reduce SRE toil—incident triage, runbook execution, alert enrichment, capacity and cost analysis—and own their guardrails, rollback, and reliability
- Own the platform that lets other engineers build agents safely: sandboxed execution, scoped credentials, tool/MCP integrations, cost controls, and agent observability
Legion's mission is to turn hourly jobs into good jobs. The platform is AI-driven, cloud-native, and designed to optimize labor efficiency while enhancing employee experience. The company is fully remote and globally distributed.
REQUIREMENTS:
Basic Qualifications:
- 12+ years experience in SRE, DevOps, or other SaaS operations
- 3+ years hands-on experience managing AWS services and cloud infrastructure using Terraform (AWS Certification preferred)
- 3+ years experience with containerized cloud solutions using Docker, Kubernetes, or AWS EKS; familiarity with HELM charts
- 5+ years experience with at least one programming language (Go, Python, Bash, Perl)
- 5+ years experience with infrastructure automation (Terraform, Ansible, etc.), CI/CD pipelines (GIT, Jenkins, etc.), and configuration management tools (Ansible, Chef, etc.)
- Hands-on experience with Linux/Unix platforms (RedHat, CentOS, Ubuntu, Amazon Linux)
- Hands-on experience building agentic AI systems in production—tool integrations, evaluation, and failure handling—beyond use of AI coding assistants
- Bachelor's or Master's degree in Computer Science/Engineering or related field
Other Qualifications:
- Track record of introducing AI into an engineering org's SDLC with measurable results, and judgment to identify where it should not be used
- 3+ years experience with observability tools (Splunk, Nagios, Elasticsearch, Kibana, CloudWatch, Logstash) and scaling these systems
- 3+ years experience with AWS RDS or Aurora MySQL
- Demonstrated ability to work with remote teams