Member of Technical Staff, Mid-training
Develops multimodal mid-training data pipelines for omni models handling text, image, video, and audio. Requires ML expertise, scaling laws knowledge, experiment design, and strong engineering in large-scale data frameworks like Spark and Ray.
About the job
Tech Stack
- Python
- JAX and XLA
- Spark
- Ray
Focus
- Scale synthetic coding data to trillions of tokens with large-scale docker verification.
- Distill the intelligence of flagship models into flash models through synthetic data generation.
- Optimize mid-training data mixtures to boost the ceiling for RL.
- Engineer long-context data recipes.
- Develop robust and diverse evaluation for mid-training checkpoints.
Ideal Experience
- Expertise in ML and large model scaling, with familiarity across all kinds of scaling laws.
- Strong ability to design ML experiments.
- Familiarity with state-of-the-art techniques for curating AI training data for text, image, audio, and video modalities.
- Strong engineering abilities in Spark, Ray, and other frameworks for large-scale data processing.
Annual Salary Range
$180,000 - $440,000 USD
Benefits
Base salary is just one part of our total rewards package at xAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.
Skills
Python, JAX, Xla, Spark, Ray, Machine Learning, LLMs, Scaling Laws, Data Processing, Synthetic Data
Similar jobs
ML Engineering jobsBuild and deploy agentic systems that power AI-driven creative video workflows. The role requires 5+ years of experience, production ML or agentic pipeline development, context engineering, and expertise in evaluation and agent infrastructure.
Build and advance agentic machine-learning systems for multimodal creative tasks, with a focus on video understanding, reasoning, control, and tool use. The role requires strong production ML or agent-pipeline experience and deep knowledge of modern LLM techniques.
Build evaluation methods, RL environments, agent tooling, and scalable infrastructure that make subjective qualities such as design and taste measurable for frontier AI models. The role requires experience with evaluations, RL environments, ML or post-training, plus strong backend engineering skills.
Build and scale generative video and multimodal models, optimizing training and inference for efficiency, throughput, and ultra-low latency. The role requires deep learning systems expertise, strong PyTorch/CUDA experience, and the ability to move research models into production.
Build production-grade AI agents, evaluation infrastructure, and developer tooling that make AI-assisted engineering faster, safer, and reusable across teams. The role requires software engineering experience, platform or internal developer-product experience, and hands-on expertise with LLM integration and orchestration.