Latest ML Engineering jobs
Job results
Build and lead production-grade machine learning services for Coinbase’s conversational AI ecosystem, coordinating vendor and internal LLM systems through a scalable orchestration layer. The role requires 5+ years of ML and software engineering experience, strong Python skills, and expertise in modern generative AI architectures.
Build and improve large-scale ML models for personalization and recommendation across Pinterest surfaces including Shopping. Requires 5+ years applying ML methods and experience with big data pipelines.
Senior Research Engineer building data synthesis, analysis, and management tooling for AI safety model training and evaluation. Requires strong software engineering, statistics, and ML framework expertise.
Lead technical vision and execution for post-training systems that transform foundation models into intelligent, engaging products. Drive alignment algorithms, RL, and infrastructure for large-scale LLM training and serving.
Research and develop improvements to models' personalization and agentic capabilities through reinforcement learning, dataset creation, and post-training methods. Requires strong ML engineering skills and research experience with novel models.
Own end-to-end data strategy and RL environment creation for domain-specific knowledge work (finance, healthcare, legal). Combine applied research with hands-on data sourcing, vendor management, and model performance measurement.
Lead design and delivery of high-priority AI initiatives across multiple codebases. Build and ship AI-powered features with strong backend fundamentals and product sense.
Design and implement evaluation frameworks and pipelines for AI systems using Evaluation-Driven Development. Build Python-based test suites, LLM graders, and measurement systems that guide prompt iteration and production deployment decisions.
Senior Engineer building multi-agent AI systems, LLM integrations, and backend automation services that power Marketing Operations. Owns technical direction for agentic infrastructure connecting models to business systems.
Build and own end-to-end AI agents for enterprise customers, integrating latest text/voice models and iterating based on real-world usage. Requires 8+ years of software engineering experience with Python and TypeScript.
Lead the development and deployment of agentic AI systems, LLMs, and autonomous workflows to enhance trading operations and decision-making at a global principal trading firm. Requires 3+ years in AI roles, strong Python skills, and technical leadership experience.
Research Engineer building and deploying production voice and multimodal ML models. Requires expert PyTorch, large-scale model training experience, and shipping user-facing ML systems.
Perception engineer building robot training data pipelines: hand/body pose estimation, multi-camera calibration, SLAM, perception model training, and TensorRT/CUDA deployment on embedded platforms.
Technical leader building agent infrastructure, observability, evals, and guardrails for production AI systems at Watershed. Requires 6+ years backend/platform/AI engineering experience and production TypeScript systems.
Build and scale ML infrastructure platform for autonomous vehicle development, focusing on automated resource provisioning, high-performance workload scheduling, and petabyte-scale data processing pipelines.
Build and scale ML infrastructure platform for autonomous vehicle model development, focusing on automated resource provisioning, high-performance workload scheduling, and petabyte-scale data processing pipelines.
Build and optimize ML infrastructure for autonomous vehicles, focusing on model optimization, compilers, and deployment across the autonomy stack. Requires 2+ years in ML optimization and strong Python/C++/CUDA skills.
Build and optimize ML infrastructure for autonomous vehicles, focusing on model optimization, compilers, and deployment of large models to Nuro's self-driving fleet. Requires 3+ years ML optimization experience and strong Python/C++/CUDA skills.
Build and own algorithmic systems that evaluate providers, make recommendations, and optimize healthcare outcomes for cost, quality, and access. Requires 3+ years shipping data-driven algorithms to production.
Lead optimization research applying large-scale constrained optimization and ML to real-time trading decisions. Requires 5-10+ years experience, strong math/ML background, production coding skills, and PhD-level coursework.
Optimizes end-to-end machine learning inference performance across kernels, compilers, systems, and clusters. The role requires strong computer architecture knowledge, experience with performance modeling and profiling, and proficiency in C++ and Python.
Builds and leads production-grade AI-powered conversational systems, including orchestration services connecting LLM frameworks, vendor AI, internal agents, and human workflows. Requires at least five years of machine learning and software engineering experience, strong Python skills, and expertise in modern generative AI architectures.
Build the AWS-based AI Operations platform that hosts, governs, and secures Hi Marley’s internal agents and tools. The role requires principal-level systems expertise, deep AWS and identity authorization knowledge, production LLM experience, and strong product and communication skills.
Conduct optimization research and implement large-scale constrained optimization models that drive real-time trading decisions, working across the full research lifecycle from theory to production. Requires PhD-level coursework and strong applied research background in optimization.
Senior IC building and maintaining ML underwriting and credit decisioning models for Cash App Borrow and Afterpay. Owns full modeling lifecycle including experimentation, calibration, deployment, and monitoring.
Staff AI Engineer building and deploying production AI agents, data pipelines, and ML features for a platform serving 70M+ users. Requires 6+ years experience with AI engineering, MLOps, and data-intensive systems.
Founding ML Engineer building production ML systems for governance, security, and agentic platform capabilities at Docker. Requires 5+ years applied ML experience shipping systems and 4+ years backend/infra engineering.
Build and optimize scalable machine learning systems and AI-powered customer support products, taking initiatives from concept through production. The role requires 4+ years of ML engineering experience, advanced technical expertise, and fluency in English and Mandarin.
Own the policy framework and agent harness for an AI underwriting agent. Design configurable policies, build evals, and ship production LLM/agentic systems. Partner with ML engineers and product leadership on the agentic roadmap.
Staff ML Engineer leading end-to-end identity verification ML systems including document authenticity, face matching, liveness detection, GNN-based identity graphs, and behavioral risk models. Requires 8+ years production ML experience and domain expertise in biometrics or fraud detection.
Technical lead for the shared, accelerator-agnostic inference runtime serving Claude. Owns architecture, performance, and validation for GPU/TPU/Trainium platforms in a high-scale distributed systems environment.
Design, build, and maintain LLM integrations powering AI features. Own end-to-end delivery from requirements through production monitoring with focus on scalability and reliability.
Founding Staff ML Engineer building production ML systems for governance, security, and agentic platform capabilities at Docker. Owns architecture, data pipelines, evaluation, and model lifecycle while mentoring the growing team.
Principal Engineer setting technical vision and building AI/ML infrastructure for Generative AI and Recommender Systems at Pinterest, scaling to hundreds of millions of inferences per second. Requires deep expertise in distributed systems and proven cross-org technical leadership.
Build production-grade machine learning and LLM-based solutions that automate support processes and augment technical teams. The role requires at least three years of software engineering experience, machine learning experience, and proficiency with Java or Python services.
8-month modelling residency embedded in production and research projects, focusing on adaptive algorithms, real-time learning, model efficiency, and cross-stack optimization. Requires Python, deep learning frameworks, and ML optimization experience.
Research Engineer advancing Claude's code generation capabilities through reinforcement learning. Design RL environments, build verifiers, run training experiments on frontier models, and improve training pipelines for real software engineering tasks.
3-month research fellowship for early-career researchers working on frontier Multimodal LLMs, generative modeling, and real-time audiovisual AI. Own a research problem in pretraining, post-training, RL, evaluation, or multimodal modeling. Strong PyTorch and first-author tier-1 paper required.
Build and fine-tune vision-centric VLMs and generative models using Pinterest's visual-text datasets. Requires 2+ years industry computer vision experience and an M.S. or Ph.D.
Build and scale Snowflake's Cortex Training LLM post-training platform, handling distributed GPU scheduling, orchestration, and productionizing research for enterprise-scale model adaptation.
New/recent PhD to own RL and post-training for large-scale omni models. Build and scale the full RL/post-training stack including rollout, optimization, reward modeling, and evaluation for real-time audiovisual AI.
Early-career engineer optimizing inference for real-time multimodal AI avatars. Focus on KV cache strategies, serving frameworks, quantization, and latency reduction for LLMs and diffusion models.
Lead projects building and deploying large-scale ASR/NLP/LLM systems for meeting intelligence. Architect training, fine-tuning, and inference pipelines using PyTorch/JAX and own ML systems from research to production.
Develops and optimizes computer vision algorithms for aerial object detection and tracking in defense systems. Requires Master's/PhD, 2+ years experience in CV/deep learning, Python/C++/OpenCV/PyTorch expertise.
Designs and deploys end-to-end ML pipelines and MLOps practices for AI-driven counter-drone systems in defense applications. Requires 5+ years ML engineering experience, Python/C++ proficiency, cloud/containerization, and computer vision expertise.
Staff ML Engineer setting technical direction for autonomous mineral refining using reinforcement learning and simulation. Owns modeling, validation, and deployment of control systems on live industrial equipment.
Build and deploy reinforcement learning models to autonomously control mineral refining facilities, optimizing recovery rates, energy use, and uptime in real operating plants.
Build and optimize AI agent orchestration and reasoning systems for the insurance industry. Requires 6+ years in ML/AI, strong Python skills, and exceptional LLM prompting ability.
Build and operate production multi-agent AI systems that turn longitudinal health data into actionable clinical recommendations. Requires 6+ years building production ML or backend systems plus 1+ years building agentic AI.
Leads development and deployment of machine learning, perception, and planning systems for defense-focused autonomous vehicles. Requires 5+ years of production software experience, strong C++ or Python skills, deep learning expertise, and eligibility for UK Security Clearance.