Member of Technical Staff, Agent Code
Researches and engineers code-generating LLMs and autonomous agent systems for enterprise automation. The role requires deep code-model expertise, strong Python and deep-learning framework skills, scalable systems experience, and a PhD with relevant publications.
About the job
Responsibilities
- Stay up to date with research in code LLMs, agents, and related fields, implementing novel ideas into systems.
- Design and implement scalable strategies to train code models and deploy agent frameworks for inference and sampling.
- Collaborate with the pretraining team to create supervised fine-tuning trajectories and work on existing and new reinforcement learning algorithms.
- Improve existing benchmarks and design new benchmarks reflecting enterprise user needs.
- Lead experiments on state-of-the-art compute infrastructure for frontier LLMs.
Requirements
- PhD in Computer Science, Machine Learning, or a related field.
- Publications in top-tier venues such as NeurIPS, ICML, ICLR, ACL, or EMNLP.
- Deep expertise in code LLMs and agent systems, including active contributions to code model development.
- Hands-on experience with frontier LLMs and applications in code generation or automation.
- Strong software engineering skills and proficiency in Python and PyTorch, TensorFlow, or similar frameworks.
- Experience with distributed systems, cloud infrastructure, and scalable architectures.
- Proactive, self-motivated approach and passion for ambitious, open-ended problems.
Benefits and Compensation
- Competitive compensation and equity.
- Weekly lunch stipend of $75/£75 or equivalent in local currency.
- Comprehensive health and dental benefits, including a separate mental health budget.
- RRSP matching, 401(k), or pension scheme, depending on location.
- Up to six months of fully topped-up parental leave for either parent.
- Annual enrichment benefits for arts and culture, fitness and wellness, quality time, and workspace improvements.
- Education and learning stipend for conferences, courses, and coaching.
- Six weeks of paid vacation (30 working days).
- Travel budget for remote employees to visit other offices and an annual company offsite.
- Coworking benefit for employees not near an office.
- $500 home office stipend.
Skills
Python, PyTorch, TensorFlow, LLMs, Code Generation, Autonomous Agents, Reinforcement Learning, Supervised Fine-Tuning, Distributed Systems, Cloud Infrastructure, Scalable Architectures, Benchmarking
Similar jobs
AI Research jobsDesigns, validates, and publishes rigorous evaluations and benchmarks for frontier AI systems across agentic, coding, safety, and expert-domain applications. The role requires strong research publication experience, scientific writing, experimental rigor, and the ability to deliver reproducible evaluation systems.
Develop and optimize production-scale LLM inference systems across distributed runtimes, GPU kernels, scheduling, and model-system co-design. The role requires a bachelor’s degree and at least five years of experience in inference, distributed AI, GPU systems, or high-performance computing.
Conduct applied research on AI agents, designing experiments and evaluation systems to improve reliability, context retention, and multi-step task completion. The role requires strong AI/ML research, engineering, experimental design, and communication skills.
Research Engineer building large-scale AI capability evaluations, telemetry, data pipelines, and analysis tools for Anthropic’s Takeoff Intel team. The role requires hands-on large language model experimentation, rapid prototyping, data expertise, and strong research collaboration.
Conduct applied research on foundation models for fraud detection using large-scale behavioral and financial-risk data. The role spans experimentation, evaluation, production deployment, and cross-functional work on model governance, requiring 4+ years of applied ML experience and strong Python and SQL skills.