Researcher developing evaluations, red-teaming pipelines, and novel mitigations for frontier AI safety risks. The role requires deep technical expertise, research engineering experience, advanced training in computer science or machine learning, and proficiency in Python or similar languages.
295k – 445k/yr
On-site4+ YOEAI Research
About the role
Responsibilities
Identify emerging AI safety risks and develop methodologies to explore and mitigate their impact.
Build and continuously refine evaluations for assessing frontier AI risks, collaborating with internal and external domain experts.
Set research directions and strategies to make AI systems safer, more aligned, and more robust.
Contribute to best-practice guidelines for AI safety within OpenAI and across the industry.
Evaluate and design red-teaming pipelines to examine the end-to-end robustness of safety systems and identify areas for improvement.
Develop novel safety mitigations using techniques from interpretability, control, alignment, and related domains.
Collaborate cross-functionally with experts in misalignment, cybersecurity, biology, and other fields to develop scalable and enforceable safety systems.
Requirements
2+ years of experience in AI safety, particularly in areas such as RLHF, human-AI collaboration, interpretability, or control.
Ph.D. or another advanced degree in computer science, machine learning, or a related field.
Experience working with large-scale AI systems.
4+ years of research engineering experience.
Proficiency in Python or similar programming languages.
Enthusiasm for long-term AI safety and technical approaches to safe AGI.
Ability to apply methods from interpretability, robustness, alignment, and control to improve model safety.
Strong technical depth and ability to collaborate closely across functions.
Compensation
Annual salary range: $295,000–$445,000.
Skills
ai safetyPythonRLHFinterpretabilityhuman-ai collaborationrobustnessalignmentcontrolred teamingMachine LearningCybersecuritylarge-scale ai systems
Conduct strategic, technically rigorous research to anticipate and mitigate loss-of-control risks from increasingly capable AI systems, including recursive self-improvement. The role combines hypothesis-driven research, rapid prototyping, safety evaluations, monitoring, and institutionalizing effective interventions.
295k – 445k/yrOn-siteAI Research
Research Engineer / Research Scientist / AI Systems Engineer, RSI
OpenAISan Francisco, CA
Develop AI systems that automate and accelerate research by designing evaluations, building research agents and orchestration infrastructure, and improving model capabilities through training and synthetic data. The role suits strong research or engineering generalists experienced with LLMs, evaluations, agents, infrastructure, or distributed systems.
295k – 445k/yrHybridAI Research
Research Engineer / Research Scientist - Personal AGI, Personalization
OpenAISan Francisco, CA
Researches and develops memory and personalization improvements for frontier models through post-training, reinforcement learning, dataset creation, and evaluations. The role requires strong machine-learning expertise, research craftsmanship, and the ability to work across a large codebase with research and product teams.
295k – 555k/yrHybridAI Research
Researcher, Multimodal Safety
OpenAISan Francisco, CA
Conducts research to improve the safety of multimodal AI systems spanning text, vision, and audio. The role requires experience building multimodal models, post-training frontier systems, designing safety evaluations, and translating research findings into reliable model behavior.
295k – 445k/yrHybridAI Research
Researcher, Synthetic RL
OpenAISan Francisco, CA
Develops novel reinforcement learning techniques using synthetic environments and feedback to enhance large-scale AI models. Designs experiments, analyzes dynamics, and integrates research into production systems; requires strong RL/ML background and engineering skills.