Skip to content
AnthropicAnthropic

Research Engineer, Performance RL

Research Engineer on the Code RL team advancing AI models' ability to write efficient code for accelerators. Requires deep expertise in accelerators like CUDA/ROCm and ML frameworks like JAX/PyTorch, plus experience across kernels, model code, and distributed systems.

About the job

Responsibilities

  • Invent, design and implement RL environments and evaluations.
  • Conduct experiments and shape our research roadmap.
  • Deliver your work into training runs.
  • Collaborate with other researchers, engineers, and performance engineering specialists across and outside Anthropic.

Requirements

  • Expertise with accelerators (CUDA, ROCm, Triton, Pallas), ML framework programming (JAX or PyTorch).
  • Worked across the stack – kernels, model code, distributed systems.
  • Know how to balance research exploration with engineering implementation.
  • Passionate about AI's potential and committed to developing safe and beneficial systems.

Nice-to-Haves

  • Experience with reinforcement learning.
  • Experience porting ML workloads between different types of accelerators.
  • Familiarity with LLM training methodologies.

Compensation

Annual Salary: $350,000—$850,000 USD

Skills

CUDA, Rocm, Triton, Pallas, JAX, PyTorch, Reinforcement Learning, Distributed Systems, Llm Training, Kernels

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Research Software Engineer, Post Training
$350k+/yrHybridML Engineering

Build and operate the engineering systems that support post-training research, including reinforcement learning infrastructure, sandboxed execution, data pipelines, and agent scaffolding. The role requires strong Python and systems engineering skills, project ownership, and a relevant bachelor’s degree or equivalent experience.

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

AI Infrastructure Engineer
$350k+/yrOn-site4+ YOEML Engineering

Operates and improves the infrastructure powering large-scale post-training and reinforcement learning runs, partnering with researchers to debug failures, improve reliability, and automate recovery. Requires 4+ years operating distributed production systems and strong Python, Go, or C++ skills.

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Research, General Agents
$350k+/yrHybridML Engineering

Research-focused engineer advancing agentic model capabilities across synthetic data, task environments, evaluations, training, and usability improvements. Requires strong Python engineering, deep learning framework experience, scalable distributed training skills, and scientific experimentation ability.

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Research, RL Scaling
$350k+/yrHybridML Engineering

Researcher focused on scaling reinforcement learning for frontier models, with ownership spanning asynchronous RL algorithms, inference and distributed training systems, and large-scale empirical studies. Requires strong Python and deep learning experience, scalable systems debugging, and rigorous research judgment.

OpenAI

OpenAI

San Francisco, CA

Machine Learning Engineer, Multimodal Perception and Authentication
$342k+/yrHybridML Engineering

Develop multimodal perception and authentication systems combining visual, audio, and other sensor signals for real-world AI products. The role requires machine learning expertise, practical research-to-system experience, and proficiency in Python and PyTorch with comfort in C++.