Software Engineer, ML Infrastructure, Optimization
Build and optimize ML infrastructure for autonomous vehicles, focusing on model optimization, compilers, and deployment across the autonomy stack. Requires 2+ years in ML optimization and strong Python/C++/CUDA skills.
About the job
About the Work
- Optimize Nuro’s autonomy stack with cutting-edge optimization techniques like quantization, low precision inference, and model pruning.
- Work with autonomy engineers to optimize, validate, and deploy large language models.
- Develop and maintain a world-class model compiler framework, FTL.
- Write robust, high-quality software to increase our confidence in our vehicle’s ability to navigate safely on-road.
- Collaborate closely with machine learning domain experts and engineers across behavior, perception and mapping to design and implement end-to-end learned ML solutions.
About You
- 2+ years of relevant experience in ML optimization infrastructure.
- Experience with ML optimization techniques such as quantization and pruning, and ML compilers.
- Experience maintaining, profiling, and optimizing GPU ML compilers & runtimes.
- Proficient in Python and working experience with C++ and CUDA.
- Working experience deep learning frameworks (like PyTorch, Jax, Tensorflow, Keras).
- Proficient in Python and working experience with C++.
- You are passionate about accelerating the benefits of robotics for everyday life.
Compensation and Benefits
- Base pay range: $160,360 - $240,540
- Annual performance bonus
- Equity
- Competitive benefits package
Skills
Python, C++, CUDA, PyTorch, JAX, TensorFlow, Keras, Quantization, Model Pruning, Ml Compilers
Similar jobs
ML Engineering jobsBuild and maintain machine learning infrastructure for autonomy teams, including model pipelines, observability, inference serving, and compiler platforms. The role requires a relevant degree, at least one year of experience, strong Python skills, and familiarity with C++.
Build and operate the infrastructure powering large-scale machine-learning training for autonomous-driving systems. The role requires Python proficiency, Kubernetes production experience, distributed-systems expertise, and ownership of reliability, observability, and operational maturity.
Build, deploy, and optimize AI applications and machine learning models for customer use cases while contributing to an internal ML platform. This new graduate role requires a technical master’s degree, hands-on ML or LLM experience, and strong customer communication skills.
Build and ship production algorithmic systems that improve healthcare quality, access, and cost outcomes. The role combines machine learning, optimization, experimentation, and LLM productionization, requiring at least two years of relevant industry or advanced-degree experience.
Build replayable enterprise environments, evaluation systems, graders, and post-training workflows for AI agents. The role spans machine-learning research and production engineering and requires 1–7 years of software or ML systems experience.