SlipstreamJobsFresh Startup & VC-Backed Jobs

Research Engineer, Generative Video

Captions - New York, NY, United States - In-office - posted 2026-09-28

Apply on the company site

SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.

Mirage is an AI-native video platform that uses natural language to intelligently orchestrate production and editing. The platform leverages contextual awareness to execute creative decisions a professional editor would make, improving productivity for experienced teams while making video creation accessible to anyone. As a Research Engineer on the Generative Video team, you'll work at the intersection of research and systems engineering, focusing on making advanced video generation models faster, more efficient, and capable of ultra-low latency, real-time generation. You'll tackle foundational problems in generative media that remain largely unsolved across the industry. Key Responsibilities: - Train and optimize large-scale video and multimodal models - Improve efficiency across training and inference (memory, latency, cost) - Implement techniques such as distillation, quantization, and pruning to aggressively accelerate diffusion and autoregressive generation - Build and maintain distributed training systems - Optimize GPU utilization, parallelism, and throughput - Develop tooling for experimentation, evaluation, and debugging - Translate research models into robust, production-ready systems - Monitor and improve model performance in real-world usage The company is well-funded (raised $75M, backed by Index Ventures, Kleiner Perkins, Sequoia Capital, Andreessen Horowitz, General Catalyst, and notable founders including Kevin Systrom and Mike Krieger) and recognized as one of Fast Company's Most Innovative Companies 2025 and Forbes AI 50. All roles require in-person presence at the NYC HQ in Union Square. Requirements: - BS/MS/PhD in CS, ML, or related field - 2+ years of professional industry experience - Strong experience in deep learning systems and infrastructure - Expertise in PyTorch, CUDA, Triton, and distributed training (FSDP, etc.) - Experience scaling and optimizing large models under low-latency inference constraints - Strong debugging and performance profiling skills - Ability to move quickly from prototype to production

Similar roles