Skip to content
MirageMirageNew York, NY

Research Engineer, Generative Video

Build and scale generative video and multimodal models, optimizing training and inference for low latency, efficiency, and production reliability. The role requires deep learning systems expertise, strong PyTorch/CUDA experience, and at least two years of professional industry experience.

175k – 275k/yr
On-site5+ YOEML Engineering

About the role

Responsibilities

  • Train and optimize large-scale video and multimodal models.
  • Improve efficiency across training and inference, including memory, latency, and cost.
  • Implement distillation, quantization, and pruning techniques to accelerate diffusion and autoregressive generation.
  • Build and maintain distributed training systems.
  • Optimize GPU utilization, parallelism, and throughput.
  • Develop tooling for experimentation, evaluation, and debugging.
  • Translate research models into robust, production-ready systems.
  • Monitor and improve model performance in real-world usage.

Requirements

  • BS, MS, or PhD in Computer Science, Machine Learning, or a related field.
  • 2+ years of professional industry experience.
  • Strong experience with deep learning systems and infrastructure.
  • Expertise in PyTorch, CUDA, Triton, and distributed training, including FSDP.
  • Experience scaling and optimizing large models under low-latency inference constraints.
  • Strong debugging and performance-profiling skills.
  • Ability to move quickly from prototype to production.

Benefits

  • Comprehensive medical, dental, and vision plans.
  • 401(k) with employer match.
  • Commuter benefits.
  • Catered lunch multiple days per week.
  • Dinner stipend when working late.
  • Grubhub subscription.
  • Health and wellness perks.
  • Multiple team offsites per year and monthly team events.
  • Generous PTO policy.

Skills

Machine LearningDeep LearningPyTorchCUDAtritonDistributed Trainingfsdpdiffusion modelsautoregressive modelsmodel quantizationmodel pruninggpu optimizationperformance profilingmultimodal modelsinference optimization

Similar roles

ML Engineering jobs
Mirage

Research Engineer, Agentic Systems

MirageNew York, NY

Build multimodal agentic systems that use large language models for creative video analysis and editing. The role spans model training, structured generation, evaluation, failure analysis, and deployment of production ML pipelines.

175k – 275k/yrOn-siteML Engineering
xAI

Software Engineer

xAIPalo Alto, CA

Build and own a mission-critical research platform for evaluating AI model capabilities and behaviors at xAI. Design instruments, datasets, grading schemes, and infrastructure to measure, diagnose, and improve models while shipping delightful internal tools.

175k – 275k/yrOn-site3+ YOEML Engineering
Mirage

Software Engineer, ML Products

MirageNew York, NY

Build and ship end-to-end agentic systems and architectures for creative video workflows at an AI-native video platform. Requires 5+ years building production ML/agentic pipelines, deep RAG/context engineering experience, and strong evaluation skills.

175k – 275k/yrOn-site5+ YOEML Engineering
Applied Intuition

Robotic Software Engineer, Perception

Applied IntuitionSunnyvale, CA +1

Develop and integrate real-time AI/ML perception algorithms and sensor fusion software for autonomous vehicles across land, air, sea, and space domains. Requires MS/PhD or 5+ years experience with multi-modal sensors, ML deployment, and Linux/Docker; US citizenship and security clearance eligibility mandatory.

175k – 250k/yrOn-site5+ YOEML Engineering
Taste Labs

AI Engineer

Taste LabsSan Francisco, CA

Build AI systems for taste evaluation, synthetic data, agent tooling, retrieval, crawling, inference, and user-facing experiences. The role favors hands-on experience shipping LLM and agent systems, early-stage startup experience, and comfort inventing infrastructure in ambiguous domains.

175k – 275k/yrOn-siteML Engineering