SlipstreamJobsFresh Startup & VC-Backed Jobs

Staff Software Engineer, Billing

Docker - Remote - Remote - posted 2026-10-01

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Salary: USD 170,350 - 275,550 / annual

Docker is seeking a Staff Software Engineer to own the infrastructure supporting the Billing Platform team. Docker is a globally distributed, remote-first company trusted by 20+ million monthly users and 20+ billion container image pulls. The company is at the center of AI-driven software development, providing sandboxed environments, verified images, and secure infrastructure for autonomous workflows. The Billing Platform Engineering team owns the systems that power Docker's commercial model. In this role, you will own and evolve infrastructure supporting billing platform services across compute, storage, networking, CI/CD, and observability. Key responsibilities include: - Design and maintain Infrastructure as Code (Terraform) for billing system infrastructure on AWS, establishing module patterns and standards for the team - Build and own observability systems—metrics, logging, alerting—with focus on billing accuracy and payment reliability - Define deployment patterns and runbooks optimized for AI-agent-assisted development workflows, including clear rollback procedures, safe promotion gates, and automated validation - Partner with software engineers on service design, bringing infrastructure constraints and operational requirements into conversations before code is written - Identify systemic risks and drive improvements spanning team or organizational boundaries - Lead incident response for billing system issues; participate in on-call rotation including evenings, weekends, and holidays as needed - Mentor engineers across the team, raising technical standards organization-wide You will ship code in your first week through Docker's agent-first development workflow. Within 30 days, you will shadow on-call and understand the system deeply. By 90 days, you will own infrastructure components and deliver measurable improvements from design to production. Within one year, you will be the team's trusted authority on billing infrastructure and help define what AI-agent-assisted infrastructure operations look like at scale. REQUIREMENTS: - 8+ years in platform, infrastructure, or SRE roles supporting production SaaS systems at scale - Deep AWS expertise: ECS or EKS, RDS (Postgres preferred), networking, IAM, cost management—with experience operating these systems under real load and real incidents - Expert-level Terraform with experience designing reusable module patterns and setting standards - Experience building and owning observability stacks (Datadog, Grafana, or similar) at organizational level - Strong familiarity with CI/CD systems (Jenkins, GitHub Actions, or equivalent) including pipeline design and developer experience ownership - Kubernetes at operational and architectural level - Track record identifying systemic risks and driving improvements spanning team or organizational boundaries - Security-first mindset: threat modeling, blast radius analysis, least-privilege by default, audit trails as design requirement - Strong written English; at Staff level, written communication scales influence across teams - Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience - Meaningful plus: experience with billing, payments, or financial systems infrastructure

Similar roles