Software Engineer, ML Research
Builds distributed training, inference, and data systems for frontier coding models, working with researchers to enable fast iteration. Requires strong infrastructure background and intuitions about language models.
About the job
What you might do
- Build our distributed training, inference, and RL infrastructure
- Write libraries to simplify how researchers do large-scale data jobs
- Architect the systems that turn Cursor user data into effective training data
You might be a fit if
- You have a strong infrastructure/distributed systems background
- You are able to architect and ship end-to-end with high ownership
- You have strong intuitions about how language models work
- You’re excited to learn more about ML
Skills
Distributed Systems, Reinforcement Learning, PyTorch, TensorFlow, Kubernetes, ML Infrastructure, Data Pipelines, Language Models
Similar jobs
AI Research jobsConduct applied research on AI agents, designing experiments and evaluation systems to improve reliability, context retention, and multi-step task completion. The role requires strong AI/ML research, engineering, experimental design, and communication skills.
Research Engineer building large-scale AI capability evaluations, telemetry, data pipelines, and analysis tools for Anthropic’s Takeoff Intel team. The role requires hands-on large language model experimentation, rapid prototyping, data expertise, and strong research collaboration.
Conduct applied research on foundation models for fraud detection using large-scale behavioral and financial-risk data. The role spans experimentation, evaluation, production deployment, and cross-functional work on model governance, requiring 4+ years of applied ML experience and strong Python and SQL skills.
Researcher or engineer focused on designing, evaluating, and productionizing oversight systems and safety mitigations for autonomous AI agents. The role requires strong systems or security reasoning, threat-modeling ability, and experience building practical evaluations and controls.
Researcher focused on training and evaluating frontier AI agents, mining incidents, and building scalable safety measurement systems. The role requires strong research or ML engineering execution, quantitative judgment, and the ability to own ambiguous projects end to end.