SlipstreamJobsFresh Startup & VC-Backed Jobs

Senior Site Reliability Engineer, NetBox Delivery

NetBox Labs - Remote - Remote - posted 2026-09-28

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

NetBox Labs is hiring a Senior Site Reliability Engineer to join the newly formed NetBox Delivery team within the Applications group. This is a foundational role where you'll help shape how the team operates from the ground up. NetBox is a network source-of-truth platform serving thousands of companies through three distribution channels: open-source NetBox OSS, NetBox Cloud (SaaS), and NetBox Enterprise (self-managed). The Delivery team owns the entire pipeline from NetBox Core releases through to healthy, running instances in production across Cloud and Enterprise. You'll be responsible for shipping software, monitoring production health, and fixing issues at the source rather than applying workarounds. Key responsibilities include: - Owning the NetBox build and release pipeline, from base images to downstream availability on Cloud and Enterprise platforms - Building the release handoff process between NetBox Core and downstream teams to ensure new releases reach customers quickly and predictably - Improving NetBox performance and reliability in production, from application startup to Postgres query optimization - Implementing real observability for both the application and release pipeline, including monitoring, alerting, and SLO definition - Serving as the escalation point for performance and reliability issues, driving fixes back to NetBox Core when appropriate - Strengthening supply chain security and supporting SOC 2 compliance for the build pipeline - Sharing on-call duties and leading incident response and postmortems for your area The role requires hands-on technical depth across a modern cloud-native stack and the ability to drive cross-functional initiatives. REQUIREMENTS: - 5+ years in software engineering, platform engineering, or SRE with proven experience writing robust, maintainable code - Production experience with Django and Postgres at scale, including schema design, migration risk assessment, and query performance optimization under real load - Strong container build skills, including base image design, Python dependency management, and supply chain security practices (vulnerability scanning, image signing) - Hands-on experience with AWS (EC2, VPC, IAM, RDS), Kubernetes and Helm, GitHub Actions, ArgoCD or FluxCD, Terraform, and Prometheus and Grafana - Hands-on experience building inside an AI-augmented development harness, including Claude Code and workflows that make agentic tooling reliable - Track record of driving work across team boundaries, from RFC writing to executing cross-team migrations NICE TO HAVES: - Familiarity with NetBox ecosystem or network automation - Open source contributions or project involvement - Experience in B2B software startups or high-growth organizations - Deep experience with supply chain security tooling (cosign, Sigstore, SLSA) - Experience operating high-throughput or performance-sensitive systems for large enterprise customers

Similar roles