Software Engineer, RL Data
Build the tasks, rewards, environments, and data systems used to train and evaluate coding agents. The role requires strong software engineering fundamentals and experience with infrastructure, data, or distributed systems; reinforcement learning experience is a plus.
About the job
Responsibilities
- Design task sets that teach specific coding-agent capabilities and iterate using traces and evaluations.
- Analyze agent traces to identify failure modes and unexpected behaviors.
- Build systems that surface recurring agent behaviors and failures.
- Convert one-off recipes into reusable tools and processes, including improved rewards, cleaner environments, and higher-quality data.
- Partner with research to evaluate whether datasets teach the intended capabilities.
Requirements
- Strong software engineering fundamentals.
- Ability to write careful, fast code.
- Ability to decompose ambiguous capabilities into concrete, measurable tasks.
- Background in infrastructure, data, or distributed systems.
- Comfort analyzing messy real-world agent behavior and turning it into datasets or tools.
Nice-to-have
- Reinforcement learning experience.
Skills
Software Engineering, Reinforcement Learning, Distributed Systems, Data Infrastructure, Agent Evaluation, Dataset Design, Reward Design, Python
Similar jobs
ML Engineering jobsBuild and operate the engineering systems that support post-training research, including reinforcement learning infrastructure, sandboxed execution, data pipelines, and agent scaffolding. The role requires strong Python and systems engineering skills, project ownership, and a relevant bachelor’s degree or equivalent experience.
Build research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.
Build production AI capabilities for automated slide and document generation, working across LLM applications, data analysis, and content generation. The role requires 3+ years in machine learning and NLP, advanced Python, and experience with LLM frameworks and production systems.
Build and operate large-scale ranking and retrieval systems that power search relevance, including hybrid lexical/vector search, embeddings, query understanding, and permission-aware retrieval. Requires a bachelor's degree and 5+ years of ML engineering experience in ranking or information retrieval.
Develop and deploy machine learning models for biomedical research and AI products, collaborating with scientific, engineering, and product teams. Requires an advanced quantitative degree, substantial ML experience, Python proficiency, and experience bringing models into production or research applications.