SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Rapyd is a global fintech platform unifying payments, payouts, and financial services across 8+ international offices. The NOC Team Leader role bridges traditional Network Operations Center (NOC) functions with modern Site Reliability Engineering (SRE) practices, leading a 24/7 global operations team.
You will direct operational leadership for a distributed NOC team, managing incident response, service uptime, and operational excellence across production systems and infrastructure. Key responsibilities include implementing and managing monitoring and observability tools (Prometheus, Grafana, Datadog) to track system health via golden signals—latency, traffic, errors, and saturation. You'll drive automation initiatives to reduce operational toil and manual troubleshooting, improving system reliability and reducing human error.
Incident management is central to the role: you'll oversee critical incident response, ensure timely stakeholder communication, and lead post-mortem root cause analysis (RCA) to prevent recurrence. You'll coach and mentor NOC engineers and SREs, fostering a high-performance culture of continuous improvement and technical excellence.
Collaboration across engineering, development, and IT teams is essential to align system performance with business goals and SLA requirements. The role reflects a modern transformation of traditional NOCs into automated, resilience-focused SRE-driven environments. This is a hybrid position based in Tel Aviv, part of Rapyd's global operations footprint.