SlipstreamJobsFresh Startup & VC-Backed Jobs

Cloud Platform Engineer

SingleStore - Remote - Remote - posted 2026-09-17

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

SingleStore is seeking a Cloud Platform Engineer to join its engineering organization and help drive the next generation of infrastructure and platform capabilities. You will be responsible for designing, operating, and optimizing cloud environments, maintaining cloud infrastructure, and creating internal developer platforms that support fast, reliable, and secure software delivery. You will work closely with engineering, QA, data, and product teams to ensure systems are scalable, resilient, and easy to use. In this role, you will automate infrastructure provisioning, configuration management, monitoring, and operational workflows using Infrastructure as Code and scripting languages. You will own the deployment, maintenance, and lifecycle management of systems supporting engineering, including Kubernetes clusters, container registries, artifact systems, and internal developer platforms, leveraging deep expertise in Kubernetes, container runtimes, and the broader cloud-native ecosystem (Helm, Kustomize, operators). You will troubleshoot complex infrastructure and application issues, driving root-cause analysis and developing long-term remediation solutions. You will design, build, and maintain cloud infrastructure across major cloud providers (AWS, GCP, Azure), and develop/support deployments of applications, services, and monitoring with a strong focus on scalability, reliability, and cost optimization. You will develop, enhance, and maintain CI/CD pipelines using modern DevOps tooling (GitHub Actions, ArgoCD, Terraform, etc.), and develop internal tooling and automation using Terraform, Python, Go, or similar languages to streamline operational tasks and improve developer productivity. You will implement and manage security best practices across cloud environments, including identity management, secrets handling, audit logging, and network controls. Additionally, you will leverage AI/ML tools to automate repetitive DevOps tasks and operational workflows. SingleStore is a venture-backed, cloud-native database company headquartered in San Francisco with offices globally including Sunnyvale, Raleigh, Seattle, Boston, London, Lisbon, Bangalore, Dublin, and Kyiv. The company delivers a distributed SQL database that unifies transactions and analytics, enabling digital leaders to deliver exceptional, real-time data experiences. **REQUIREMENTS:** - 4-8+ years of experience in DevOps, Cloud Engineering, Systems Administration, or similar infrastructure-focused roles - Familiar with GIT and comfortable with at least one systems programming language (Golang) and one scripting language (Python, Bash) - Proficiency with Infrastructure as Code (Terraform, Pulumi, or CloudFormation) and configuration management (Ansible, Chef, or SaltStack) - Working knowledge of at least one observability stack (Prometheus, Grafana, ELK/OpenSearch, Datadog, etc.) and operational troubleshooting - Strong hands-on experience with Kubernetes management in production environments - Experience designing, developing, and/or troubleshooting distributed systems - Comfort with shells on *nix family systems - B.S. degree or equivalent experience in Engineering, Computer Science, or a related field **PREFERRED QUALIFICATIONS:** - Certifications from cloud service providers (e.g., AWS DevOps, GCP DevOps Engineer) - Strong software development skills with experience building production-grade services, automation, APIs, and infrastructure tooling - Proficiency in Go, Python, or other modern programming languages - Strong cloud networking fundamentals (VPC/VNet, subnets, NAT Gateways, Transit Gateways, routing, DNS, load balancing, private connectivity) - Hands-on experience with VPN and secure networking technologies (Tailscale, Headscale, WireGuard, or similar) - Strong Kubernetes experience including cluster architecture, networking, service discovery, ingress, observability, and troubleshooting - Understanding of distributed systems concepts and experience with highly available, scalable, and fault-tolerant systems - Natural curiosity and strong drive to learn new and adjacent technologies - Experience applying AI/ML tools (GitHub Copilot, cloud AI services, LLM-based automation, anomaly detection tools) to streamline DevOps and infrastructure workflows - Experience with LangChain, AI agents, and MCP (Model Context Protocol), including building and integrating agent-based solutions - Hands-on experience implementing AI-enhanced threat detection and security systems

Similar roles