# Machine Learning Research Scientist / Research Engineer, Post-Training

**Company:** [Scale AI](https://hotfix.jobs/companies/scale-ai)
**Location:** San Francisco, CA, Seattle, WA, New York, NY
**Role:** AI Research
**Salary:** $181k – $226k/yr
**Skills:** LLMs, Supervised Fine-Tuning, RLHF, Reward Modeling, Preference Modeling, Instruction Tuning, Deep Learning, Reinforcement Learning, Model Fine-Tuning, Data Curation, Model Evaluation, Multimodal Models, Python, Machine Learning
**Posted:** 2026-08-26

> Research novel post-training methods for large language models, focusing on preference optimization, data curation, evaluation, alignment, and robustness across text and multimodal systems. Requires advanced academic training and experience with deep learning, reinforcement learning, and post-training techniques.

## Job Description

## Responsibilities
- Research and develop novel LLM post-training techniques, including supervised fine-tuning (SFT), reinforcement learning from human feedback (RLHF), and reward modeling.
- Optimize data curation and evaluation methods for text and multimodal models.
- Design and experiment with approaches to preference optimization.
- Analyze model behavior, identify weaknesses, and propose solutions for bias mitigation and model robustness.
- Collaborate with researchers and engineers to define best practices for data-driven AI development.
- Partner with foundation model labs to provide technical and strategic input on generative AI model development.
- Publish research findings in top-tier AI conferences.

## Requirements
- Ph.D. or master's degree in Computer Science, Machine Learning, AI, or a related field.
- Deep understanding of deep learning, reinforcement learning, and large-scale model fine-tuning.
- Experience with post-training techniques such as RLHF, preference modeling, or instruction tuning.
- Published machine learning research in major conferences or journals, such as NeurIPS, ICML, ICLR, ACL, EMNLP, or CVPR.
- Excellent written and verbal communication skills.

## Nice-to-haves
- Previous experience in a customer-facing role.

## Compensation
- Base salary range: **$180,600–$225,750 USD**.
- Eligible roles may include equity compensation and benefits such as health, dental, and vision coverage, retirement benefits, a learning and development stipend, generous paid time off, and potentially a commuter stipend.

## Similar jobs

- [Machine Learning Research Scientist, Evaluations](https://hotfix.jobs/jobs/72d88855-3e74-43ec-82cd-9eaa9e95df69) - Scale AI - San Francisco, CA - $181k – $226k/yr
- [Software Engineer (Gen AI)](https://hotfix.jobs/jobs/7af58525-5b44-47cc-a964-4a0eee5a3506) - Earnin - Mountain View, CA - $181k – $222k/yr
- [Software Engineer, Applied AI Research](https://hotfix.jobs/jobs/b270e535-825d-433e-a8e9-ea9a8e8d2836) - Hightouch - Remote - $180k – $400k/yr
- [Research Engineer](https://hotfix.jobs/jobs/996c82c8-13bd-4ba2-91e1-a50a6499eed0) - Greptile - San Francisco, CA - $180k – $280k/yr
- [Research Scientist](https://hotfix.jobs/jobs/f8da4cc2-a216-4a79-b247-5d2bb83ac27e) - Counsel Health - New York, NY - $165k – $220k/yr

**Apply:** https://hotfix.jobs/jobs/aec3f24b-dccd-4095-b46f-79ea3c89be63
**Canonical:** https://hotfix.jobs/jobs/aec3f24b-dccd-4095-b46f-79ea3c89be63