SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
OnePay is a consumer fintech platform offering banking, high-yield savings, credit cards, point-of-sale lending, investing, and crypto services. Backed by Walmart and Ribbit Capital, the company serves millions of Americans and partners with employers and gig platforms to deliver embedded financial services.
As Site Reliability Engineering Manager, you will lead a small, senior SRE team while remaining deeply hands-on. This is a builder's leadership role where you are measured equally by what you and your team ship and the operating model you establish. You will report to the Head of Platform and Data Engineering.
Key responsibilities include:
- Leading a senior SRE team with focus on planning, clarity, and career growth, directing them toward high-leverage reliability work over manual ticket-driven response
- Writing production code and building automation, tooling, and golden paths yourself; contributing to codebase, conducting design and code reviews, and setting technical standards by example
- Defining service ownership, alerting, dashboards, runbooks, and deploy/rollback safety; running regular service-health reviews for supported systems
- Running incident command when needed, driving verified remediation, preventing recurrence, and building self-service incident tooling and runbook automation
- Designing durable on-call coverage through staffing, handoffs, and automation rather than open-ended volunteer hours
- Partnering across engineering teams to raise reliability standards while maintaining clear boundaries between SRE enablement and service-team ownership
The role demands someone who still codes today and is comfortable being assessed on coding quality. You will work in a fast-moving, mission-driven culture where urgency and execution are valued.
Requirements:
- Strong software engineering foundation with production-quality code in modern languages such as Python or Go
- Real reliability depth for production systems and services at scale, ideally in consumer, fintech, or other high-availability environments
- Background from a high-bar, high-scale engineering organization with strong engineering culture
- Recent people leadership experience managing a small engineering team within the last few years, with desire to keep leading while staying technical
- Hands-on expertise with observability, alerting, safe deploys, automation, and incident command (you build these, not just specify them)
- Calm, credible communication with team, partner teams, and during live incidents