SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Wrapbook is an AI-powered production finance platform trusted by Netflix, Paramount, and other major studios. The company is backed by Andreessen Horowitz, Bessemer Venture Partners, and WndrCo, with a team of 350+ building systems that help finance teams manage payroll, spend, and accounting for feature films, TV, and commercials.
As a Senior Platform Engineer I, you will own critical infrastructure and platform systems that enable Wrapbook to deliver a stable, secure, reliable, and performant product. You'll work cross-functionally with product, security, and analytics engineers to strengthen internal developer and data platforms.
Key responsibilities include:
- Owning observability, logging, and alerting for Kubernetes clusters and critical workloads
- Building and maintaining automation for Kubernetes lifecycle management (provisioning, scaling, upgrades)
- Leading infrastructure cost-optimization initiatives (right-sizing, workload scheduling, waste identification)
- Partnering with engineering teams to shape infrastructure decisions from design through deployment
- Identifying and fixing reliability bottlenecks before they become incidents
- Participating in on-call rotation to resolve service disruptions
- Leading incident response, root-cause analysis, and blameless postmortems
You'll be a hands-on contributor who drives platform reliability and enables other engineering teams to move faster.
REQUIREMENTS:
- 5+ years in platform, infrastructure, or SRE roles, including running Kubernetes in production at scale and handling day-two operations (troubleshooting, upgrades, scaling)
- Familiarity with a major cloud provider (AWS, GCP, or Azure)
- Hands-on experience with Kubernetes cluster autoscaling and workload scheduling systems (Karpenter, Cluster AutoScaler)
- Proficiency in infrastructure-as-code tools (Terraform, Crossplane, Pulumi, or equivalent)
- Hands-on experience with observability frameworks (OpenTelemetry, Prometheus, eBPF)
- Familiarity with GitOps-based delivery and tools (Argo CD, Kargo)
- Proficient in managing and deploying Kubernetes applications (Kustomize, Helm)
- Solid knowledge of Linux systems including low-level fundamentals
- Strong grasp of web and network protocols (HTTP, TLS, DNS)
- Strong programming skills in at least one modern language beyond scripting (Go, Python, Ruby)
- Comfortable using AI as a real collaborator in your work
NICE TO HAVES:
- Experience with multi-region deployments on AWS or other cloud providers
- Deep knowledge of Kubernetes internals and ecosystem (operators, CNIs, service meshes)
- Exposure to or experience with operationalizing canary infrastructure