SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Solidus Labs is a fintech company providing trade surveillance technology for detecting market manipulation, financial crime, and fraud across traditional assets, prediction markets, and crypto. The company has 20+ years of Wall Street experience and serves financial institutions and regulators globally, monitoring over a trillion events daily. Headquarters in NYC with offices in Singapore, Tel Aviv, and London.
You will own the reliability, stability, and operational support of production systems as a DevOps/SRE engineer. This is a production-focused role requiring hands-on expertise in cloud-native infrastructure and incident response.
Key responsibilities include:
- Own reliability, availability, and performance of production environments
- Operate production Kubernetes (EKS) clusters, including upgrades and Helm deployments
- Manage scaling and capacity using KEDA, Karpenter, and HPA
- Administer AWS Cloud infrastructure (EC2, Lambda, AWS Batch, Elasticache, RDS)
- Evolve infrastructure-as-code using Terraform and Helm with security best practices
- Support GitLab CI/CD pipelines and resolve deployment issues
- Design observability systems using Prometheus, Grafana, and EFK
- Troubleshoot networking issues (TLS, load balancing, VPCs, NAT, VPN)
- Support compliance and security incident response
- Lead incident response end-to-end with RCA and corrective actions
- Participate in on-call rotations
- Leverage AI-powered tools for automation and productivity
The role requires experienced DevOps/SRE expertise with demonstrated production ownership and incident response leadership.