SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Smallest.ai is seeking a hands-on DevOps and Site Reliability Engineer to build and operate production infrastructure systems. This is a deeply execution-focused role where you will work on production systems daily—deploying, scaling, observing, fixing, and improving them continuously. At Smallest, infrastructure is not a support function; it is the product.
You will take approved system architecture and translate High-Level Designs (HLD) into practical, battle-tested Low-Level Designs (LLD). The company values reliability earned through ownership, automation, and clean engineering rather than titles or years of experience.
Key responsibilities include:
- Implement and manage AWS-centric cloud infrastructure using Terraform
- Operate Kubernetes (EKS) clusters across multiple environments
- Build and maintain CI/CD pipelines using GitHub Actions
- Deploy services using Helm and Argo CD (GitOps)
- Implement canary, blue/green, and rolling deployments
- Build and manage Docker images and registries
- Configure monitoring, alerting, and logging using New Relic and CloudWatch
- Manage AWS networking: VPCs, subnets, routing, ALB/NLB, security groups
- Support RabbitMQ, Amazon SQS, Redis, and MongoDB infrastructure
- Support frontend delivery using CloudFront and AWS Amplify
- Write automation scripts in Bash, Python, or Go
- Troubleshoot incidents and participate in postmortems
The company values high ownership and accountability, comfort working in ambiguity, automation-first thinking, reliability over velocity without safety, and a strong bias toward learning from failures.
REQUIREMENTS:
- Strong hands-on experience with AWS and Kubernetes
- Solid Linux and networking fundamentals
- CI/CD pipeline design and ownership mindset
- Production debugging and incident handling experience
- Knowledge of SLOs, SLIs, and error budgets
STRONG PLUS:
- Startup or scale-up production exposure
- DevOps/SRE side projects or homelabs
- Open-source contributions in cloud-native ecosystem
- Experience with cost optimization and capacity planning
- Terraform-based infrastructure automation
- Helm and GitOps-based deployment workflows