SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: EUR 90,000 - 120,000 / annual
Hive is an operations platform for independent commerce, founded in 2020 and backed by Tiger Global, Earlybird, and Picus Capital. The company is scaling across Europe with offices in Berlin, Paris, Milan, Madrid, London, and Amsterdam.
You will join as a Senior Site Reliability Engineer to help power Hive's infrastructure, ensure platform reliability, and drive operational excellence. The role is hands-on and user-focused, working closely with engineers, product teams, analysts, and operations teams.
Key responsibilities include:
• Own production reliability: manage SLOs, alerting, and observability across metrics, logs, and traces; participate in paid on-call rotation; drive lasting fixes through post-incident reviews.
• Evolve infrastructure: design, build, and optimize Kubernetes and public cloud platforms managed as code via GitOps, with focus on performance, cost efficiency, and self-service for engineering teams.
• Manage data layer: own performance, capacity, replication, and upgrades of PostgreSQL fleet; maintain pipelines feeding analytical backends.
• Build AI as platform capability: create AI templates and coding standards for AI agents; design data access models enabling teams to move from prototype to production safely; hands-on building, not advisory.
• Connect platform and product: embed reliability into how teams ship; make targeted changes in Rails/React codebases when needed.
• Harden security: implement least-privilege access across cloud and Kubernetes, manage production access controls and secrets, remediate infrastructure and container vulnerabilities.
The company emphasizes a culture of trust, collaboration, empowerment, and constructive feedback. Team members have backgrounds from McKinsey, Amazon, Shopify, Google, Flink, Blackstone, J.P. Morgan, and DHL.
Benefits include 30 vacation days annually with sabbatical opportunity after three years, flexible working hours, choice of hardware and operating system, monthly wellness and productivity budget, virtual employee stock options, and on-call compensation on top of salary.
REQUIREMENTS:
Must-have:
• Hands-on production experience with AWS, Kubernetes (ideally Amazon EKS), and infrastructure as code (Terraform)
• PostgreSQL production operations: query and index tuning, replication, upgrades, backup and recovery
• End-to-end reliability ownership: observability (Prometheus, Grafana, Loki or similar), SLOs, incident response, on-call experience
• Real coding ability in Python, Ruby, TypeScript, or similar; build tools and services, not just glue scripts
• Daily use of AI coding tools; shipped something LLM-backed beyond a chat window
Nice to have:
• Ruby on Rails
• GitOps (Argo CD or similar)
• Data pipelines and cloud data warehouses
• Deeper cloud security experience (IAM, network segmentation, vulnerability management)