Skip to content

Latest ML Engineering jobs at Nuance Labs

6 jobs

Job results

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff - Research Fellow
$200k+/yrOn-siteML Engineering

3-month research fellowship for early-career researchers working on frontier Multimodal LLMs, generative modeling, and real-time audiovisual AI. Own a research problem in pretraining, post-training, RL, evaluation, or multimodal modeling. Strong PyTorch and first-author tier-1 paper required.

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff — RL Research
$250k+/yrOn-siteML Engineering

New/recent PhD to own RL and post-training for large-scale omni models. Build and scale the full RL/post-training stack including rollout, optimization, reward modeling, and evaluation for real-time audiovisual AI.

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff — Model Optimization and Inference
$200k+/yrOn-siteML Engineering

Early-career engineer optimizing inference for real-time multimodal AI avatars. Focus on KV cache strategies, serving frameworks, quantization, and latency reduction for LLMs and diffusion models.

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff — Model Optimization and Inference
$250k+/yrOn-site7+ YOEML Engineering

Optimize inference for real-time multimodal AI avatars. Specialize in LLM and diffusion model serving, KV cache strategies, quantization, and low-latency frameworks like vLLM and TensorRT-LLM.

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff — RL Research
$300k+/yrOn-site7+ YOEML Engineering

Own RL and post-training infrastructure for omni foundation models. Build and scale rollout, reward, and policy systems from 0→1 for real-time audiovisual AI.

Nuance Labs

Nuance Labs

Seattle, WA

Member of Technical Staff — Pretraining Infra
$300k+/yrOn-site7+ YOEML Engineering

Own and scale the distributed training infrastructure for large-scale omni model pretraining across GPU clusters, covering job orchestration, parallelism, GPU communication, data loading, and performance optimization.