SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Revolut is a fintech company on a mission to give people more from their money through spending, saving, investing, exchanging, and travel products. With 80+ million customers and 13,000+ employees globally, the company is scaling rapidly.
The Technology team builds the systems and infrastructure that power Revolut's platform. We are seeking a Software Engineer to join our Observability Platform team, which is responsible for creating a platform that continuously monitors thousands of applications, databases, and systems across the organization.
In this role, you will design, implement, and assemble scalable and resilient observability solutions across logs, metrics, and traces. You will build robust APIs and data pipelines to ingest, process, and expose observability data to product teams. Working closely with engineering teams, you will understand their observability needs and integrate solutions that empower them to monitor, alert, and debug their components effectively.
Key responsibilities include:
- Designing and implementing scalable observability solutions leveraging existing market solutions or building from scratch
- Building APIs and data pipelines for high-throughput, real-time data ingestion and processing
- Collaborating with product teams to understand observability requirements and deliver integrated solutions
- Optimizing observability infrastructure for performance, accuracy, cost-effectiveness, and user experience
- Developing tooling to automate onboarding and sunsetting of components and streamline data collection
- Contributing to the strategic roadmap of the observability platform
Your work will ensure applications are continuously optimized, incidents are quickly resolved, and friction is reduced in onboarding, normal usage, and sunsetting of observability components.
REQUIREMENTS:
- 7+ years of experience as a software engineer, with 3+ years focused on building and maintaining observability platforms or highly distributed systems
- Familiarity with monitoring, alerting, and incident response best practices
- Expertise in designing and implementing APIs and data pipelines for high-throughput, real-time data ingestion
- Practical understanding of distributed systems and their unique observability challenges
- Hands-on experience with core observability tools such as Prometheus, Grafana, Loki, ELK stack (Elasticsearch, Logstash, Kibana), Jaeger, and OpenTelemetry
- Experience with containerization and orchestration technologies (Docker, Kubernetes) and infrastructure as code tools (Ansible, Terraform)
- Proficiency in Python as primary engineering language
NICE TO HAVE:
- Previous experience in DevOps, SRE, or developer experience roles
- Experience with multiple cloud platforms (AWS, GCP, Azure) and their native observability services
- Contributions to open-source observability projects
- Track record of prototyping and sketching solutions to complex problems