AI Researcher (Multimodal Audio/Video Generation)
Leads research in multimodal audio-visual avatar generation for conversational AI, focusing on diffusion models, long-video synthesis, and integrating verbal/non-verbal signals. Requires PhD, 2-3+ years in generative models, PyTorch expertise, and top-tier publications.
About the job
Responsibilities
- Lead research efforts on audio-visual generation for avatars (Neural Avatars, Talking-Heads), with a focus on conversational settings.
- Design models that are coupled with conversation flow — capturing and generating verbal + non-verbal signals in sync.
- Drive innovation in diffusion models, long-video generation, and audio-visual modeling.
- Translate research into production by partnering with Applied ML and engineering.
- Mentor researchers, set research directions, and publish impactful work.
Requirements
- PhD or equivalent research experience, plus 2–3+ years of hands-on experience applying generative models at scale.
- Expertise in diffusion models and awareness of the latest efficiency techniques.
- Experience in multimodal generation — spanning video, audio, and language.
- Proven innovation in long-video generation and/or audio generation.
- Excellent programming skills — fluent in PyTorch and GPU-optimized workflows.
- Track record of publications in top-tier venues (CVPR, NeurIPS, BMVC, ICASSP, etc.).
- Experience leading research activities or mentoring teams.
Nice-to-Haves
- Skills in 3D graphics, Gaussian splatting, or large-scale training setups.
- Broad exposure to generative AI models beyond your specialty.
- Familiarity with software development best practices.
Skills
PyTorch, Diffusion Models, Multimodal Generation, Video Generation, Audio Generation, 3D Graphics, Gaussian Splatting, Generative AI, Long-Video Generation, Gpu-Optimized Workflows
Similar jobs
AI Research jobsSummer 2027 internship applying computer science, mathematics, statistics, and machine learning research to practical product capabilities. The intern will prototype, evaluate, and communicate solutions involving data privacy, security, governance, algorithms, and text analytics.
Conduct foundational research on LLMs and multimodal systems, designing architectures and training methods and helping move prototypes into production. The role targets PhD researchers graduating by December 2026 with strong machine-learning research and programming experience.
AI Research Intern researching agentic AI applications for customer-facing products and developing working prototypes. The role requires current pursuit of a technical bachelor's degree, prior software engineering or substantial project experience, and interest in LLMs or generative AI.
Paid internship for quantitative students contributing to AI-powered software creation, systems optimization, and developer tooling. Candidates should demonstrate strong mathematical ability, curiosity about AI, and autonomous cross-functional collaboration.
Researcher developing and publishing mechanistic interpretability techniques, building infrastructure to study model internals, and guiding alignment-focused research. Requires research experience in machine learning or a related field, strong engineering skills, and proficiency in Python or similar languages.