Researcher, Alignment
Designs and implements experiments for AI alignment research, focusing on scalable solutions for human intent following, risk evaluation, model robustness, and oversight methods in complex scenarios. Requires PhD-level research experience and strong ML engineering skills.
About the job
Responsibilities
- Develop and evaluate alignment capabilities that are subjective, context-dependent, and hard to measure.
- Design evaluations to reliably measure risks and alignment with human intent and values.
- Build tools and evaluations to study and test model robustness in different situations.
- Design experiments to understand laws for how alignment scales as a function of compute, data, lengths of context and action, as well as resources of adversaries.
- Design and evaluate new Human-AI-interaction paradigms and scalable oversight methods that redefine how humans interact with, understand, and supervise our models.
- Train models to be calibrated on correctness and risk.
- Design novel approaches for using AI in alignment research.
Requirements
- Team player – willing to do a variety of tasks that move the team forward.
- PhD or equivalent experience in research in computer science, computational science, data science, cognitive science, or similar fields.
- Strong engineering skills, particularly in designing and optimizing large-scale machine learning systems (e.g., PyTorch).
- Deep understanding of the science behind alignment algorithms and techniques.
- Can develop data visualization or data collection interfaces (e.g., TypeScript, Python).
- Enjoy fast-paced, collaborative, and cutting-edge research environments.
- Want to focus on developing AI models that are trustworthy, safe, and reliable, especially in high-stakes scenarios.
Skills
PyTorch, TypeScript, Python, Machine Learning, Alignment Research, Scalable Oversight, Human-Ai Interaction, Data Visualization
Similar jobs
AI Research jobsConducts causal inference research for financial market prediction and portfolio optimization, developing and validating models from research through live trading. Requires Ph.D.-level coursework, strong causal inference and statistics expertise, mathematical ability, and production Python skills.
Applied research scientists develop deep-learning and generative media systems for video, audio, and multimodal editing features that ship to millions of users. The role requires strong PyTorch or TensorFlow skills, rapid experimentation, and evidence of impactful research or production machine-learning work.
Build and ship agentic AI product experiences, internal automation, and customer-facing features across the stack. The role requires 5+ years of software engineering experience, hands-on experience with AI or LLM-powered products, Python proficiency, and strong autonomy.
Conducts frontier AI research for health, developing and evaluating scalable training methods, models, and agents that improve medical reasoning, reliability, and real-world outcomes. Requires exceptional machine learning or biomedical AI research depth, hands-on coding and experimentation, and end-to-end ownership of ambiguous problems.
Leads the design, measurement, publication, and adoption of APEX benchmarks evaluating frontier models on economically valuable professional work. The role requires rigorous research judgment, strong coding and statistical skills, and excellent communication across technical, commercial, and research audiences.