Research Scientist running controlled SFT and RL experiments to prove the impact of curated datasets on foundation model behavior, generalization, and alignment. Ideal for strong undergrad or master's researchers with deep obsession over data-driven model improvements.
250k – 450k/yr
On-siteAI Research
About the role
Responsibilities
Run controlled SFT and RL experiments to measure the impact of our datasets on model performance.
Help build public evals and new data types that push the frontier.
Publish external-facing research, blog posts, and technical reports.
Work with internal SPLs to iterate on data quality based on your results.
Requirements
Strong familiarity with LLM training and evaluation methodologies.
Genuine obsession with how data structure, selection, and quality drive model behavior.
Ability to design lightweight experiments, move fast, and extract actionable insights from messy results.
Comfort working across domains (finance, software engineering, policy, and more).
A bias toward building over theorizing.
Great candidates are undergrad research or master's research (but haven't done a PhD).
Compensation
Annual target cash compensation of $250-450K + meaningful equity.
Comprehensive benefits (UberEats and ride share stipend, comped Equinox, 401K with match, health, dental, and vision insurance).
Researcher designing and running experiments on chain-of-thought monitorability in frontier LLMs to support scalable oversight and alignment. Requires strong empirical ML expertise with LLMs, deep interest in model behavior/alignment/interpretability, and ability to translate ambiguous questions into concrete experiments.
250k – 445k/yr
HybridAI Research
Simulation Researcher/Engineer
Luma AILos Angeles, CA +2
As a Simulation Researcher/Engineer, you will design and build simulation environments for training general-purpose robot policies. This role involves working with generative models and classical physics simulation, developing differentiable pipelines, and driving asset generation.
250k – 450k/yr
HybridAI Research
Research Scientist - World Model
Luma AILos Angeles, CA +2
As a Research Scientist on the World Models team, you will invent next-generation world model architectures with a focus on controllability and physical consistency, develop controllability mechanisms, and define and own metrics for physical fidelity and action-following.
250k – 450k/yr
HybridAI Research
Researcher, Alignment Training
OpenAISan Francisco, CA
Senior researcher studies how training choices shape aligned behavior in frontier models, developing synthetic data, evaluation loops, and experiments to ensure durable, robust tendencies like honest reasoning and instruction-following.
250k – 445k/yr
On-siteAI Research
Researcher, Alignment Science
OpenAISan Francisco, CA
Designs and runs experiments to improve AI model intent alignment, honesty, calibration, and robustness using RL and empirical ML methods. Trains/evaluates large models like LLMs and integrates techniques into production workflows.