Research, Post-Training

350k – 475kSan Francisco, CAML EngineeringOnsiteMay 4

Summary

Develops and tunes post-training recipes for AI models, iterates on evaluations, debugs configurations, scales methodologies, and publishes research to advance collaborative intelligence. Requires Python proficiency, deep learning frameworks, and strong ML fundamentals.

About the role

What You’ll Do

Develop and tune the recipe: iterate on post-training recipes, consisting of a collection of datasets, training stages, and hyperparameters. Measure how recipe choices affect various metrics.
Iterate on evals: post-training involves a never-ending loop of defining a set of evaluations, optimizing them, and then realizing your existing evals don’t capture what matters. You’ll be responsible for both making numbers go up, and making sure the numbers are meaningful.
Debug and understand: while tuning the details of a training configuration, we often observe results that don’t quite make sense. You’ll be responsible for both getting things to work, and developing a deeper understanding, which we can bring to the next problem.
Scale and explore: post-training will involve a combination of scaling the existing methodologies and developing new ones. We’ll want to both measure how performance metrics scale with dataset size, and explore using a completely different kind of training dataset.
Publish and present research that moves the entire community forward. Share code, datasets, and insights that accelerate progress across industry and academia.

Skills and Qualifications

Minimum qualifications:

Proficiency in Python and familiarity with at least one deep learning framework (e.g., PyTorch, TensorFlow, or JAX). Comfortable with debugging distributed training and writing code that scales.
Bachelor’s degree or equivalent experience in Computer Science, Machine Learning, Physics, Mathematics, or a related discipline with strong theoretical and empirical grounding.
Clarity in communication, an ability to explain complex technical concepts in writing.

Preferred qualifications:

A strong grasp of probability, statistics, and ML fundamentals. You can look at experimental data and distinguish between real effects, noise, and bugs.
Prior experience with RLHF, RLAIF, preference modeling, or reward learning for large models.
Experience managing or analyzing human data collection campaigns or large-scale annotation workflows.
Research or engineering contributions in alignment, data-centric AI, or human-AI collaboration.
PhD in Computer Science, Machine Learning, Physics, Mathematics, or a related discipline with strong theoretical and empirical grounding; or, equivalent industry research experience.

Logistics

Compensation: Depending on background, skills and experience, the expected annual salary range for this position is $350,000 - $475,000 USD.

Skills

PythonPyTorchTensorFlowJAXMachine LearningRLHFRLAIFDistributed TrainingProbabilityStatistics

Similar roles at this salary range

All ML Engineering jobs →

OpenAI

Jun 25

Research Engineer/Research Scientist

Research Engineer/Scientist improving model capabilities for personalized AI experiences. Focus on tool-use, instruction following, evaluations, and training improvements. Requires strong ML engineering and research experience.

295k – 555kSan Francisco, CAML EngineeringHybrid7+ YOEPythonResearch

xAI

Jun 24

Member of Technical Staff

Hands-on technical contributor focused on stabilizing and advancing large language model training, fine-tuning, and research in AI/deep learning. Requires a bachelor's degree and 2+ years of experience with distributed systems, ML infrastructure, and programming in Rust/C++/Python.

324k – 396kPalo Alto, CAML EngineeringOn-site2+ YOEC++GPU

xAI

Jun 24

Member of Technical Staff

Hands-on technical leader building and scaling large language models and AI systems. Requires 3-5+ years of AI/ML experience with strong Python and deep learning frameworks.

324k – 396kPalo Alto, CAML EngineeringOn-site5+ YOEC++JAX

Anthropic

Jun 23

Research Engineer, Safeguards Labs

Research engineer on the Safeguards Labs team building and evaluating novel safety methods to detect misuse, strengthen model safeguards, and reduce real-world harm from Claude.

350k – 850kSan Francisco, CA +1ML EngineeringHybridPythonClassifiers

OpenAI

Jun 19

Research Engineer / Research Scientist

Research and develop improvements to models' personalization and agentic capabilities through reinforcement learning, dataset creation, and post-training methods. Requires strong ML engineering skills and research experience with novel models.

295k – 555kSan Francisco, CAML EngineeringHybrid7+ YOEPythonPyTorch

Apply