Member of Technical Staff - Voice Model
Develop voice AI models for natural, low-latency spoken interactions on the Grok team. Handle data pipelines, model training with JAX/PyTorch, evaluations, and product integrations. Requires Python expertise, large-scale data processing, and distributed systems experience.
About the job
Responsibilities
- Design and execute large-scale speech data curation and processing pipelines, including collection of diverse real-world audio, synthetic data generation, and automated annotation workflows.
- Work on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques.
- Build and iterate a comprehensive evaluation framework covering objective metrics, human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure.
- Work closely with product teams to integrate voice models into applications and real-time environments, define spoken interaction specifications, and handle the full lifecycle from prototype to global-scale deployment.
Basic Qualifications
- Python expert with deep proficiency in writing clean, efficient code for AI/ML systems.
- Hands-on experience processing large-scale datasets using tools like Spark and Ray for cleaning, augmentation, and feature extraction.
- Proficiency in pre-training and post-training speech-language models using JAX/PyTorch, including supervised fine-tuning, reinforcement learning, and optimizations for accuracy, factuality, natural spoken style, detail, and multilingual fluency.
- Ability to set up and run rigorous evaluation pipelines: objective metrics, human preference studies, content factuality checks, and iterative A/B testing.
- Experience building or working with large-scale distributed training and inference systems on Kubernetes.
- Proactive, self-driven attitude — ready to grind in a fast-paced, high-caliber team.
Compensation and Benefits
$150,000 - $450,000 USD base salary, plus equity, comprehensive medical, vision, dental coverage, 401(k), short & long-term disability insurance, life insurance, and various perks.
Skills
Python, Spark, Ray, JAX, PyTorch, Kubernetes, Supervised Fine-Tuning, Reinforcement Learning, Speech-Language Models, Data Curation
Similar jobs
ML Engineering jobsBuild and teach reliable AI agent systems through customer workshops, technical content, guidance, and reference implementations. The role requires strong Python and agent-development experience plus a background delivering customer-facing technical training.
Build reinforcement-learning environments, evaluations, datasets, and scalable infrastructure for frontier AI capabilities. The role suits a high-agency generalist engineer with experience in agents, evaluations, or RL workflows and strong communication skills.
Develop and productionize machine- and deep-learning algorithms for biosignal and EEG data used in medical devices, clinical development, and diagnostics. The role requires 4+ years of industry experience, DSP and statistics expertise, PyTorch proficiency, and familiarity with regulated environments and production ML practices.
Develop and deploy ML-first behavior prediction and planning systems for autonomous vehicles, forecasting the motion and interactions of road users. Requires a bachelor's degree, deep learning lifecycle expertise, and at least three years of production software experience with C++ or Python.
Build the AI platform behind fab2, including model infrastructure, agent systems, evaluations, and tools for engineering and fab operations. The role requires strong production software engineering skills and comfort working across frontend, backend, infrastructure, and data.