SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Twelve Labs is building AI infrastructure to enable machines to understand video at scale. Video comprises 90% of global data, yet most remains inaccessible to machine understanding. The company has raised over $300M from top-tier investors including NEA, Radical Ventures, Amazon, NVIDIA, Snowflake, and Databricks, with advisors including Fei-Fei Li and Silvio Savarese. Operating globally from San Francisco, Seoul, New York, and London, core R&D happens in Seoul.
We seek a Senior Platform Engineer to design, build, and operate the systems underpinning TwelveLabs' AI platform. This is a highly hands-on role requiring deep expertise in infrastructure, backend systems, and reliability engineering. You will own significant portions of production systems, moving beyond ticket handoffs to directly investigate incidents, validate assumptions, trace systems end-to-end, identify root causes, and drive solutions to production.
You will design and build scalable infrastructure and internal platforms supporting the AI product; own reliability and operational health from incident response through root cause analysis and long-term improvements; build and operate CI/CD pipelines and infrastructure automation using tools like Terraform; improve production reliability, performance, scalability, and cost efficiency; develop platform tooling enabling faster, safer product and engineering team deployments; collaborate closely with backend, ML, researcher, and product teams to understand requirements and architect solutions; and lead technical investigations on ambiguous problems from diagnosis through production rollout.
Ideal candidates have production experience with AWS, GCP, or Azure; deep expertise in Linux, networking, containers, and Kubernetes; hands-on Infrastructure as Code experience (Terraform, Ansible); CI/CD and deployment system design experience; strong observability, monitoring, logging, and incident response skills; backend programming in Python, Go, TypeScript, or similar languages; experience debugging complex multi-layer production issues; solid distributed systems and cloud architecture knowledge; and a bias toward root cause analysis over symptom mitigation.
Preferred: 7+ years in infrastructure, platform engineering, SRE, or DevOps; experience building developer or infrastructure platforms used by multiple teams; large-scale B2B SaaS operations; security and compliance-focused system design; backend service and distributed systems development; thoughtful trade-off analysis between performance, reliability, and cost; startup or high-ownership environment experience; and English communication ability.
Benefits include hybrid work with autonomy and collaboration, latest MacBooks and ~$700 home office stipend (refreshed every 3 years), unlimited LLM tokens for tech staff, ~$1,400 annual professional development budget, English training and global buddy programs, evening/weekend taxi support, ~$7,200 annual corporate card for meals and transport, office snacks and evening meals after 7pm, annual health checkups for employee and one family member, group insurance options, flu shot coverage, and 2-week paid holiday break in December.