SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Systems Reliability Engineer (SRE), Edge

Cloudflare - London, United Kingdom - Hybrid - posted 2026-08-28

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Cloudflare is seeking a Senior Systems Reliability Engineer to build and operate its Edge platform, which runs in over 320 cities across 120+ countries. This role sits at the intersection of systems, network, and software engineering, focusing on automation, scalability, and operational excellence. You will own a wide portfolio of applications and services, working to improve service availability, performance, and operational velocity across Cloudflare's global network. The team operates on a "follow the sun" model with offices in East Asia, Europe, and North America, supporting services 24/7/365. You'll develop tools and systems that make infrastructure failure-resistant and ready to scale, leveraging monitoring, alerting, and diagnostics tools while enhancing platform capabilities. Key responsibilities include identifying and owning technical problems, collaborating across teams to solve them, and driving an "automate everything" philosophy. You'll work with distributed systems at massive scale, optimizing for reliability and performance. The role requires on-call flexibility outside standard working hours to address technical issues. Required qualifications: 3+ years in an SRE role or equivalent, strong Linux systems knowledge, software development skills in Go or Python, understanding of distributed systems and large-scale design tradeoffs, intermediate networking knowledge (DNS, HTTP, BGP, IP anycast). Desirable experience includes Linux kernel work, performance analysis, configuration management (Saltstack, Chef, Puppet, Ansible), load balancing/reverse proxies (Nginx, Varnish, HAProxy), SQL/time-series databases (PostgreSQL, Prometheus, Grafana), and continuous release engineering. Open source contributions and experience in 24/7 production environments are bonuses. The team uses tools including Nginx, PostgreSQL, Docker, Prometheus, Grafana, Consul, Nomad, and Salt.

Similar roles