# Researcher, Frontier Risk Mitigations

**Company:** [OpenAI](https://hotfix.jobs/companies/openai)
**Location:** San Francisco, CA
**Role:** AI Research
**Salary:** $295k – $445k/yr
**Experience:** 4+ years
**Skills:** ai safety, Python, RLHF, interpretability, human-ai collaboration, robustness, alignment, control, red teaming, Machine Learning, Cybersecurity, large-scale ai systems
**Posted:** 2026-08-10

> Researcher developing evaluations, red-teaming pipelines, and novel mitigations for frontier AI safety risks. The role requires deep technical expertise, research engineering experience, advanced training in computer science or machine learning, and proficiency in Python or similar languages.

## Job Description

## Responsibilities
- Identify emerging AI safety risks and develop methodologies to explore and mitigate their impact.
- Build and continuously refine evaluations for assessing frontier AI risks, collaborating with internal and external domain experts.
- Set research directions and strategies to make AI systems safer, more aligned, and more robust.
- Contribute to best-practice guidelines for AI safety within OpenAI and across the industry.
- Evaluate and design red-teaming pipelines to examine the end-to-end robustness of safety systems and identify areas for improvement.
- Develop novel safety mitigations using techniques from interpretability, control, alignment, and related domains.
- Collaborate cross-functionally with experts in misalignment, cybersecurity, biology, and other fields to develop scalable and enforceable safety systems.

## Requirements
- 2+ years of experience in AI safety, particularly in areas such as RLHF, human-AI collaboration, interpretability, or control.
- Ph.D. or another advanced degree in computer science, machine learning, or a related field.
- Experience working with large-scale AI systems.
- 4+ years of research engineering experience.
- Proficiency in Python or similar programming languages.
- Enthusiasm for long-term AI safety and technical approaches to safe AGI.
- Ability to apply methods from interpretability, robustness, alignment, and control to improve model safety.
- Strong technical depth and ability to collaborate closely across functions.

## Compensation
- Annual salary range: $295,000–$445,000.

## Similar roles

- [Researcher, Recursive Self-Improvement Safety](https://hotfix.jobs/jobs/1f8fd130-4760-4d3a-af28-47681ec30870) - OpenAI - San Francisco, CA - $295k – $445k/yr
- [Research Engineer / Research Scientist / AI Systems Engineer, RSI](https://hotfix.jobs/jobs/9de7d179-a863-41e6-afd6-3f2c7ab08ad0) - OpenAI - San Francisco, CA - $295k – $445k/yr
- [Research Engineer / Research Scientist - Personal AGI, Personalization](https://hotfix.jobs/jobs/8af0d507-7137-40c8-98e1-11be16336041) - OpenAI - San Francisco, CA - $295k – $555k/yr
- [Researcher, Multimodal Safety](https://hotfix.jobs/jobs/2cc84f9d-4c40-42d1-858a-f8dcb305ef5f) - OpenAI - San Francisco, CA - $295k – $445k/yr
- [Researcher, Synthetic RL](https://hotfix.jobs/jobs/1d48e4f7-9589-45b3-96ed-05276e32344e) - OpenAI - San Francisco, CA - $295k – $445k/yr

**Apply:** https://hotfix.jobs/jobs/7ff5dad3-4092-4016-af83-7697cc0a93a8
**Canonical:** https://hotfix.jobs/jobs/7ff5dad3-4092-4016-af83-7697cc0a93a8