Skip to content
TavusTavus

AI Researcher (Multimodal Audio/Video Generation)

Leads research in multimodal audio-visual avatar generation for conversational AI, focusing on diffusion models, long-video synthesis, and integrating verbal/non-verbal signals. Requires PhD, 2-3+ years in generative models, PyTorch expertise, and top-tier publications.

About the job

Responsibilities

  • Lead research efforts on audio-visual generation for avatars (Neural Avatars, Talking-Heads), with a focus on conversational settings.
  • Design models that are coupled with conversation flow — capturing and generating verbal + non-verbal signals in sync.
  • Drive innovation in diffusion models, long-video generation, and audio-visual modeling.
  • Translate research into production by partnering with Applied ML and engineering.
  • Mentor researchers, set research directions, and publish impactful work.

Requirements

  • PhD or equivalent research experience, plus 2–3+ years of hands-on experience applying generative models at scale.
  • Expertise in diffusion models and awareness of the latest efficiency techniques.
  • Experience in multimodal generation — spanning video, audio, and language.
  • Proven innovation in long-video generation and/or audio generation.
  • Excellent programming skills — fluent in PyTorch and GPU-optimized workflows.
  • Track record of publications in top-tier venues (CVPR, NeurIPS, BMVC, ICASSP, etc.).
  • Experience leading research activities or mentoring teams.

Nice-to-Haves

  • Skills in 3D graphics, Gaussian splatting, or large-scale training setups.
  • Broad exposure to generative AI models beyond your specialty.
  • Familiarity with software development best practices.

Skills

PyTorch, Diffusion Models, Multimodal Generation, Video Generation, Audio Generation, 3D Graphics, Gaussian Splatting, Generative AI, Long-Video Generation, Gpu-Optimized Workflows

Immuta

Immuta

College Park, MD

Product Research Internship
$52k+/yrHybridAI Research

Summer 2027 internship applying computer science, mathematics, statistics, and machine learning research to practical product capabilities. The intern will prototype, evaluate, and communicate solutions involving data privacy, security, governance, algorithms, and text analytics.

Fireworks AI

Fireworks AI

San Mateo, CA
Member of Technical Staff, Research
$200k+/yrOn-siteAI Research

Conduct foundational research on LLMs and multimodal systems, designing architectures and training methods and helping move prototypes into production. The role targets PhD researchers graduating by December 2026 with strong machine-learning research and programming experience.

OnePay

OnePay

United States

AI Research Intern
$67k+/yrRemoteAI Research

AI Research Intern researching agentic AI applications for customer-facing products and developing working prototypes. The role requires current pursuit of a technical bachelor's degree, prior software engineering or substantial project experience, and interest in LLMs or generative AI.

Replit

Replit

Foster City, CA

Cohort 0
No salary listedHybridAI Research

Paid internship for quantitative students contributing to AI-powered software creation, systems optimization, and developer tooling. Candidates should demonstrate strong mathematical ability, curiosity about AI, and autonomous cross-functional collaboration.

OpenAI

OpenAI

San Francisco, CA

Researcher, Alignment Interpretability
$295k+/yrOn-site2+ YOEAI Research

Researcher developing and publishing mechanistic interpretability techniques, building infrastructure to study model internals, and guiding alignment-focused research. Requires research experience in machine learning or a related field, strong engineering skills, and proficiency in Python or similar languages.