Skip to content
AnthropicAnthropic

Anthropic Fellows Program — AI Safety

Fellows conduct research on AI safety, collaborating with Anthropic researchers to develop evaluation methods and alignment techniques. Requires strong interest in AI safety, CS/math background, and ML experience.

About the job

Responsibilities

  • Conduct research on AI safety topics.
  • Collaborate with Anthropic researchers on safety evaluations and alignment techniques.
  • Develop and test methods for evaluating AI model safety.

Requirements

  • Strong interest in AI safety and alignment.
  • Background in computer science, mathematics, or related fields.
  • Experience with machine learning frameworks like PyTorch or TensorFlow.

Nice-to-Haves

  • Prior research experience in AI safety.
  • Programming skills in Python.
  • Familiarity with reinforcement learning.

Skills

Python, PyTorch, TensorFlow, Machine Learning, Reinforcement Learning, Ai Safety

Improbable

Improbable

Remote

AI Researcher
No salary listedRemoteAI Research

Conduct applied research on AI agents, designing experiments and evaluation systems to improve reliability, context retention, and multi-step task completion. The role requires strong AI/ML research, engineering, experimental design, and communication skills.

Anthropic

Anthropic

San Francisco, CA

Research Engineer, Takeoff Intel
$350k+/yrHybridAI Research

Research Engineer building large-scale AI capability evaluations, telemetry, data pipelines, and analysis tools for Anthropic’s Takeoff Intel team. The role requires hands-on large language model experimentation, rapid prototyping, data expertise, and strong research collaboration.

Sardine

Sardine

United States

Applied AI Research Scientist
No salary listedRemote4+ YOEAI Research

Conduct applied research on foundation models for fraud detection using large-scale behavioral and financial-risk data. The role spans experimentation, evaluation, production deployment, and cross-functional work on model governance, requiring 4+ years of applied ML experience and strong Python and SQL skills.

OpenAI

OpenAI

San Francisco, CA

Researcher, Agent Safety, Oversight and System Mitigations
$380k+/yrHybridAI Research

Researcher or engineer focused on designing, evaluating, and productionizing oversight systems and safety mitigations for autonomous AI agents. The role requires strong systems or security reasoning, threat-modeling ability, and experience building practical evaluations and controls.

OpenAI

OpenAI

San Francisco, CA

Researcher, Agent Safety, Training and Evaluations
$380k+/yrHybridAI Research

Researcher focused on training and evaluating frontier AI agents, mining incidents, and building scalable safety measurement systems. The role requires strong research or ML engineering execution, quantitative judgment, and the ability to own ambiguous projects end to end.