Research Engineer / Research Scientist, Vision
Research engineer/scientist building and evaluating vision capabilities for Claude models. Requires 7+ years ML/computer vision experience and work across pretraining, RL, and agentic infrastructure.
About the job
What you'll do
- Run experiments to evaluate architectural variants, data strategies, and SL and RL techniques to improve Claude’s vision
- Develop and test tools, skills, and agentic infrastructure that enable Claude to reason over visual inputs
- Create evaluations and benchmarks that measure progress on multimodal capabilities across training and deployment
- Work with our product org to find solutions to our most vexing API customer challenges related to vision and spatial reasoning
You may be a good fit if you
- Have 7+ years of ML, computer vision, and software engineering experience through industry, academia, or other projects
- Are familiar with the architecture, training, and operation of large vision language models
- Have experience creating and evaluating large synthetic and real-world visual training datasets
- Have experience engaging in systematic prompting, finetuning, or evaluation
- Are results-oriented, with a bias towards flexibility and impact
- Enjoy pair programming and cross-team collaboration
- Care about the societal impacts of your work
Strong candidates may also have experience with
- Large-scale pretraining, SL, and RL on language models
- Deep learning research on images, video, or other modalities
- Developing complex agentic systems using LLMs
- High-performance ML systems (GPUs, TPUs, JAX, PyTorch)
- Large-scale ETL and data pipeline development
Representative projects
- Running experiments to determine ideal training datamixes and parameters for a synthetically generated vision dataset
- Finetuning Claude to maximize its performance using a particular set of agent tools/skills
- Building a pipeline to ingest and process a novel source of visual training data
- Designing and running experiments to evaluate the scalability of two architectural variants
Skills
Computer Vision, Machine Learning, Large Vision Language Models, Synthetic Data Generation, Prompting, Finetuning, Evaluation, Pretraining, Supervised Learning, Reinforcement Learning, Agentic Systems, JAX, PyTorch, Gpus, Tpus
Similar jobs
ML Engineering jobsLeads and builds a team of Applied Scientists developing production algorithmic systems for healthcare optimization, LLM applications, and member engagement. Requires 6+ years of relevant industry experience, strong technical judgment, and hands-on expertise across machine learning and optimization.
Build production machine learning systems for model customization, post-training, evaluation, and AWS-native API integration. The role requires 7+ years of relevant engineering experience and expertise in deep learning, transformers, LLM fine-tuning, and production ML infrastructure.
Leads a hands-on AI engineering team developing, evaluating, and deploying large-scale multimodal and video models. The role combines post-training, inference optimization, product experimentation, technical roadmap ownership, and people management.
Leads Discord’s Safety ML team, setting technical direction and overseeing production machine learning systems for content understanding, account integrity, and platform abuse. Requires substantial machine learning and engineering management experience, hands-on technical depth, and experience delivering ML systems at scale.
Senior AI Engineer responsible for production LLM agents that enrich business identity data through web discovery, verification, classification, and risk scoring. The role requires strong asynchronous Python, agent and evaluation expertise, browser automation, and experience operating AI systems in production.