SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Augmodo is expanding its global engineering footprint and building a follow-the-sun infrastructure model. This role positions you as the anchor infrastructure engineer for the European/EMEA cohort, serving as the primary infrastructure contact during EU business hours while supporting cross-regional needs across APAC and EMEA time zones.
You will own system reliability and infrastructure leadership for the EU region, balancing active incident response and system maintenance with long-term infrastructure building. Your responsibilities include:
- Serve as the primary infrastructure contact during EU working hours, extending coverage to support APAC (Australia) and EMEA regional needs
- Build, scale, and optimize infrastructure dedicated to supporting growing EMEA teams and regional customer deployments
- Proactively manage system health, monitoring, alerting, and incident response for GCP-based workloads to maintain high availability
- Automate deployments, manage Kubernetes clusters, and build internal tooling that empowers developers to ship code safely and fast
- Establish effective async workflows and handoff protocols with other infrastructure teams to ensure seamless 24/7 continuity
The role is based in London (BST/GMT) or Central European (CET/CEST) time zones, with a schedule providing comfortable overlap with the US Bay Area team for syncs and handoffs (typically early afternoon EU / morning PT), while providing critical coverage during EU hours and the start of the APAC business day.
REQUIREMENTS:
- 5+ years of experience as a software engineer in a production environment
- Bachelor's degree in computer science, engineering, or math
- 3+ years of hands-on experience running and scaling workloads in Google Cloud Platform (GCP)
- 3+ years of experience managing production Kubernetes clusters (GKE experience is a plus)
- Proficiency in Python and Go for building infrastructure automation, APIs, and operational tooling
- Demonstrated experience with Infrastructure as Code (Terraform/Pulumi), CI/CD pipelines, and observability tools (Prometheus, Grafana, Datadog, etc.)
- High degree of self-direction and clear async communication skills, with experience thriving in distributed team environments
- Familiarity with hosting, serving, or orchestrating agentic AI systems and LLM workflows
NICE-TO-HAVE:
- Experience building and scaling modern data pipelines or stream-processing architectures
- Experience working with early-stage technical companies