Manager, RL Algorithms & Decoder
Lead a team developing large-scale RL and ML models for autonomous vehicle behavior planning and driving decisions. Requires RL/ML expertise, production ML pipeline experience, and 3+ years in leadership.
About the job
Responsibilities
- Lead, mentor, and grow a team of individual contributors, fostering innovation and continuous improvement.
- Develop and organize strategy for Onboard Behavior ML Models to generate driving plans for autonomous vehicles.
- Interface with partner teams (Onboard Perception, Cost Planner, Simulation, Validation, Data Science, Systems Engineering, QA, ML Infra) to identify model improvement opportunities.
- Set short- and long-term technical direction for the team and collaborate on company-wide directions.
- Provide technical guidance and leadership in design and development of large-scale training models, ensuring efficient inference with partner teams.
- Establish and monitor KPIs to measure effectiveness of work packages and drive continuous improvement.
- Manage resource allocation, ensuring projects are appropriately staffed with necessary tools and support.
Requirements
- Expertise with Reinforcement Learning and Machine Learning for at least one of: Planning, LLMs, VLAs/VLMs, recommendation systems.
- Extensive experience with programming and algorithm design, strong mathematics skills.
- MS or PhD degree in computer science or related field.
- 5+ years of experience with production Machine Learning pipelines, with at least 3 years in a leadership or management role.
Nice-to-Haves
- Conference or Journal publications in Machine Learning or Robotics related venues.
- Prior experience working with autonomous vehicles or robotics, diffusion models, large scale training.
Skills
Reinforcement Learning, Machine Learning, Planning, LLMs, Vlas, Vlms, Recommendation Systems, Python, Algorithm Design, Mathematics, Large Scale Training
Similar jobs
ML Engineering jobsBuild and deploy LLM-powered tools, agents, and ecosystem infrastructure with life sciences research institutions. The role requires deep scientific or biomedical research experience, production software development expertise, and the ability to translate partner workflows into scalable AI systems.
Build research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.
Build and optimize the production LLM inference runtime for frontier models on OpenAI’s custom silicon. The role spans scheduling, distributed execution, memory and KV-cache management, performance tooling, and hardware-software co-design.
Build and operate machine learning models for sales roleplay, scoring, and coaching products, owning the lifecycle from fine-tuning and evaluation through production and on-device deployment. The role emphasizes open-source models, latency and privacy optimization, and rigorous model testing.
Build and optimize OpenAI’s inference stack for AWS Trainium across high-performance kernels, compilers, runtimes, and model execution. The role requires systems programming and accelerator experience, with opportunities to solve end-to-end performance problems for frontier-scale AI models.