SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Alan is a prevention-focused health insurance company serving 1M+ members across 40K+ companies, with €800M+ ARR. The Tech Foundations area enables product teams through world-class infrastructure, developer experience, operational excellence, and security.
The Infra crew within Tech Foundations is hiring a Platform Engineer to operate and evolve the foundations Alan runs on. You'll own multi-cloud infrastructure (AWS/GCP), data platforms (PostgreSQL, events pipeline, async workloads), reliability and observability (Datadog, incident management), governance/compliance/FinOps, and AI-augmented infrastructure operations.
Key responsibilities include:
- Design and implement robust infrastructure systems with focus on reliability and observability
- Own and evolve cloud footprint, Infrastructure-as-Code (Terraform, Terramate), and architectural decisions for international expansion
- Operate and evolve data platforms and application foundations that product teams build on
- Own platform uptime and latency SLOs end-to-end; evolve observability stack and incident response processes
- Manage cross-cutting concerns: infrastructure cost (FinOps), IAM posture, backup/disaster recovery
- Integrate AI assistants into incident investigation, runbook execution, and toil reduction
2026 focus areas: multi-cloud expansion for international growth, service architecture decoupling to scale from 1M to 10M members, operational excellence improvements, and AI-augmented infrastructure operations.
The role is hands-on: you write code (Python, TypeScript, Terraform) for automation and tooling, ship to production, operate it (rollouts, alerting, on-call, incident response). You measure impact by how fast product teams ship without consulting you. You build reusable patterns, guardrails, and reliable-by-default abstractions.
Tech stack: Python/Flask, TypeScript, React, React Native; AWS, GCP, Kubernetes, PostgreSQL, Redis, Terraform/Terramate, Datadog, Cloudflare, GitHub Actions; monorepo with daily deployments and distributed ownership.
Requirements:
- 3+ years in Platform Engineering, DevOps, Site Reliability Engineering, or equivalent
- Strong experience with cloud platforms (AWS or GCP preferred)
- Comfortable writing code for automation and tooling (Python/TypeScript experience helpful but not required)
- Experience with observability, monitoring, alerting, or incident management tools and practices
- Ability to design and implement robust infrastructure systems with focus on reliability and observability
- Experience maintaining uptime targets while optimizing infrastructure costs
- Track record building and maintaining internal or external tooling (APIs, CLIs, configuration-as-code, etc.)
- Fluent in English (French not required)
- Must be legally eligible to work in France, Belgium, or Spain
- Aiming to hire at C1 level or above per Alan's engineering career path, but high potential and curiosity valued above checklist completion