Research Engineer, Generative Video
Mirage — Tracked from its ashby job board
About the role
Mirage is an AI-native video platform that intelligently orchestrates production and editing through natural language. Our models leverage contextual awareness to execute the same creative decisions a professional editor would — dramatically improving productivity for experienced teams, while making video creation accessible to anyone.
We’re an interdisciplinary team addressing some of the most difficult technical and creative challenges in generative media. As an early member of our team, you’ll tackle foundational problems that remain largely unsolved across the industry, driving an outsized impact on the future of creative expression.
More about us
Product https://mirage.app/captions (Captions by Mirage)
Research https://mirage.app/research (Our Models and Agents)
Updates https://x.com/trymirage (Mirage on X / twitter)
TechCrunch https://techcrunch.com/2026/03/24/mirage-raises-75m-to-continue-building-models-for-its-ai-video-editing-app-captions/, Forbes AI 50 https://www.forbes.com/companies/captions/?list=ai50, Fast Company https://www.fastcompany.com/91270234/video-most-innovative-companies-fast-company-2025-youtube-roku-tubi-vimeo-captions-descript-cour-procreate-synthesia-beeble (press)
Our Investors
We’re very fortunate to have some the best investors and entrepreneurs backing us, including Index Ventures, Kleiner Perkins, Sequoia Capital, Andreessen Horowitz, General Catalyst, Uncommon Projects, Kevin Systrom, Mike Krieger, Lenny Rachitsky, Antoine Martin, Julie Zhuo, Ben Rubin, Jaren Glover, SVAngel, 20VC, Ludlow Ventures, Chapter One, and more.
Please note that all of our roles will require you to be in-person at our NYC HQ (located in Union Square)
About the Role
Mirage is seeking an ML Engineer to build and scale the systems powering our video generation models. You’ll work on novel modeling approaches, training objectives, scaling strategies, and inference optimization and efficiency to bring cutting-edge models into production.
This role sits at the intersection of research and systems engineering, focusing on making advanced models faster, more efficient, and capable of ultra-low latency, real-time generation.
Responsibilities
- Train and optimize large-scale video and multimodal models
- Improve efficiency across training and inference (memory, latency, cost)
- Implement techniques such as distillation, quantization, and pruning to aggressively accelerate diffusion and autoregressive generation
- Build and maintain distributed training systems
- Optimize GPU utilization, parallelism, and throughput
- Develop tooling for experimentation, evaluation, and debugging
- Translate research models into robust, production-ready systems
- Monitor and improve model performance in real-world usage
What makes you a great fit
- BS/MS/PhD in CS, ML, or related field
- 2+ years of professional industry experience
- Strong experience in deep learning systems and infrastructure
- Expertise in PyTorch, CUDA, Triton, and distributed training (FSDP, etc.)
- Experience scaling and optimizing large models under low-latency inference constraints
- Strong debugging and performance profiling skills
- Ability to move quickly from prototype to production
BENEFITS:
- Comprehensive medical, dental, and vision plans
- 401K with employer match
- Commuter Benefits
- Catered lunch multiple days per week
- Dinner stipend every night if you're working late and want a bite!
- Grubhub subscription
- Health & Wellness Perks
- Multiple team offsites per year with team events every month
- Generous PTO policy
Captions provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Please note benefits apply to full time employees only.
What the index says about this role
- First seen by JobLarper — Aug 3, 2026, 5 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
- What Machine Learning Engineer roles ask for — across 1,108 indexed openings: ML (78%), AI/LLM (55%), Python (53%), REST/APIs (38%), Cloud (25%). This posting names ML.
- Mirage is hiring actively — 6 open roles indexed.
Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.
More open roles at Mirage
- Research Engineer, Agentic SystemsUnion Square, New York City · Mid · $175K – $275K • Offers Equity
- Software Engineer, BackendUnion Square, New York City · Mid · $175K – $275K • Offers Equity
- Software Engineer, ML SystemsUnion Square, New York City · Mid · $175K – $275K • Offers Equity
- Software Engineer, iOSUnion Square, New York City · Mid · $175K – $275K • Offers Equity
- Software Engineer, Web ProductUnion Square, New York City · Mid · $175K – $275K • Offers Equity
Similar Machine Learning Engineer roles at other companies
- Staff Machine Learning Engineer, Tools and Framework AIApple · Cupertino, United States
- Staff Machine Learning Engineer, Agentic Systems - MoveworksElement AI · Mountain View, CALIFORNIA, US
- Staff Machine Learning Engineer, Agentic App PlatformElement AI · Mountain View, CALIFORNIA, US
- Display Algorithm Engineer / Machine Learning EngineerApple · San Francisco Bay Area, United States
- Senior Clinical Research ScientistOmada Health · Remote, USA · Remote
- ML Software Research EngineerApple · Pittsburgh, United States
- AI EngineerForward Networks · Santa Clara
- PhD Research Scientist Intern - Reinforcement Learning, ImagesCanva · London, England, GB
Browse all machine learning engineer jobs in united states — 658 open roles across 272 companies.