# Researcher, Alignment Interpretability

**Company:** [OpenAI](https://hotfix.jobs/companies/openai)
**Location:** San Francisco, CA
**Role:** AI Research
**Salary:** $295k – $500k/yr
**Experience:** 2+ years
**Skills:** Mechanistic Interpretability, Ai Safety, Ai Alignment, Python, Deep Learning, Machine Learning, Neural Networks, Research Engineering, Quantitative Reasoning, Large-Scale Ai Systems
**Posted:** 2026-09-01

> Researcher developing and publishing mechanistic interpretability techniques, building infrastructure to study model internals, and guiding alignment-focused research. Requires research experience in machine learning or a related field, strong engineering skills, and proficiency in Python or similar languages.

## Job Description

## Responsibilities
- Develop and publish research on techniques for understanding representations of deep neural networks.
- Engineer infrastructure for studying model internals at scale.
- Collaborate across teams on research projects.
- Guide research directions toward demonstrable usefulness and long-term scalability.

## Requirements
- Ph.D. or research experience in computer science, machine learning, or a related field.
- 2+ years of research engineering experience.
- Proficiency in Python or similar programming languages.
- Experience or strong interest in AI safety, alignment, mechanistic interpretability, or related disciplines.
- Ability to work with large-scale AI systems.
- Strong quantitative reasoning, engineering, and research skills.

## Nice-to-haves
- Deep understanding of technical paths toward safe AGI.
- Experience studying internal representations and model behavior.

## Similar jobs

- [Member of Technical Staff](https://hotfix.jobs/jobs/6a518687-b334-4cf1-b709-06bab1bd9a0a) - Perplexity - San Francisco, CA - $220k – $405k/yr
- [Member of Technical Staff, Research](https://hotfix.jobs/jobs/40fb1099-ecd4-4081-a872-fb41cccd960d) - Fireworks AI - San Mateo, CA - $200k – $300k/yr
- [Research Engineer - New Grad](https://hotfix.jobs/jobs/c05647f9-c94b-4cdf-a7bd-f988d713ef41) - Applied Intuition - Sunnyvale, CA - $140k – $200k/yr
- [AI Research Intern: Foundation Models](https://hotfix.jobs/jobs/63788a25-96a2-4aea-b0e0-13d31a9be98d) - GrayMatter Robotics - Los Angeles, CA - $83k – $104k/yr
- [AI Research Intern](https://hotfix.jobs/jobs/fafb16d8-0cdf-40cb-8e7d-52bd9524131f) - OnePay - Remote - $67k – $67k/yr

**Apply:** https://hotfix.jobs/jobs/5c9afe98-7164-473e-86c4-f732ec6ac9be
**Canonical:** https://hotfix.jobs/jobs/5c9afe98-7164-473e-86c4-f732ec6ac9be