Latest ML Engineering jobs
Job results
Builds and productionizes AI platform components including LLMs, agentic frameworks, and core agent capabilities. Requires 3-6+ years backend experience with Python, Golang, AWS, Kubernetes, and AI-native product shipping.
Design domain-specific problems and datasets to evaluate advanced generative AI models, provide expert insights for improvement, and co-author research publications. Requires 5+ years software engineering or equivalent expertise with strong coding in Python, Java, Rust, etc.
Designs evaluation pipelines, metrics, and user studies for speech generation and recognition models (ASR/TTS). Trains evaluation models, builds dashboards, and collaborates with ML, data, and product teams to improve performance on large-scale systems.
Designs and deploys LLM-based ML systems for personalized music recommendations, playlist generation, and session experiences at Spotify scale. Requires expertise in ML, NLP, generative AI, and large-scale data processing tools.
Builds foundational, production-grade multimodal ML systems for content understanding, moderation, ranking, risk detection, and platform decisioning. The role shapes technical strategy, improves automation and reliability, and mentors engineers across content, policy, and trust and safety initiatives.
Develops advanced post-training and reinforcement learning techniques like RLHF/DPO and reward modeling to enhance AI model reasoning, truthfulness, and real-world capabilities at xAI. Seeks passionate AI enthusiasts obsessed with truth-seeking models; prior experience preferred but not required.
Forward Deployed AI Engineer works directly with customers to identify pain points, builds and deploys AI-powered solutions and automations (50% coding), and drives business impact through rapid iteration. Requires Python/JavaScript proficiency, AI tool experience, and founder-like ownership in an in-person environment.
Senior/Staff Software Engineer builds end-to-end AI products for clients, spending 75% time coding and 25% with stakeholders like CTOs. Requires 8+ years experience shipping products, applying software engineering to AI systems with high ownership.
Early career software engineer builds and integrates AI agents and solutions using frameworks like LangChain and RAG pipelines to enhance childcare platform features. Requires bachelor's in CS/engineering, Python/JS proficiency, and AI/ML familiarity; hybrid onsite in SF office 3 days/week.
Owns Lovable’s end-to-end post-training pipeline for language models, adapting reinforcement learning and preference optimization to code-generation and agent workloads. The role combines production engineering, distributed training, evaluation, deployment, and rapid experimentation.
Builds production-grade generative AI systems powering patient engagement, provider workflows, and clinical operations in mental healthcare. Requires 10+ years software engineering, deep Python/cloud expertise, and 2+ years scaling AI products with foundation models and LLM patterns.
Leads team building and scaling production ML pipelines for entity extraction and data classification across text, documents, and OCR. Requires 5+ years deploying ML systems, strong NLP expertise, Python proficiency, and translating research to production.
Lead development of deep learning models for medical image analysis (MRI, X-ray) to accelerate clinical trials, create AI biomarkers, and integrate imaging AI into drug development workflows. Requires PhD and 5+ years building production-grade medical imaging AI.
Lead development of deep learning models for medical imaging analysis to accelerate drug development and clinical trials. Requires PhD with 5+ years building production AI models for MRI/X-ray analysis.
Leads the strategy, development, and scaling of machine learning systems for content safety, policy enforcement, and compliance. The role requires production ML experience, modern frameworks, rigorous evaluation, and the ability to influence cross-functional technical direction.
Leads technical strategy and architecture for content quality ML signals at Pinterest, driving GenAI safety, signal development, and cross-team adoption in ranking and decision systems. Requires expertise in scalable ML, content modeling, and cross-functional leadership.
Optimizes ML models for performance on embedded compute platforms in ADAS/AD stacks, focusing on inference efficiency, pruning, quantization, and profiling across GPU/CPU/SoC architectures. Requires 3+ years experience with deep learning frameworks and embedded systems.
Builds and deploys AI agents to automate internal workflows at Brex by embedding with teams, integrating systems and APIs, and measuring impact. Requires 4+ years experience shipping AI/automation with LLMs, agent frameworks, and databases.
Designs and ships production AI systems including agentic workflows, RAG pipelines, and LLM integrations for an AI-native ERP platform serving finance teams. Requires 3+ years backend experience and 2+ years production AI with Python proficiency.
Models inference performance across application, model, and fleet layers using microbenchmarks to build cost-to-serve estimates. Analyzes workloads end-to-end, enhances bottleneck detection tools, and collaborates on optimizations for latency, throughput, and cost.
Designs and owns architecture for production-grade AI analyst systems, including prompts, tools, context, evaluation, and safety. Partners with teams to build reliable AI capabilities for analytical workflows, emphasizing scalability and quality.
Build AI systems for taste evaluation, synthetic data, agent tooling, retrieval, crawling, inference, and user-facing experiences. The role favors hands-on experience shipping LLM and agent systems, early-stage startup experience, and comfort inventing infrastructure in ambiguous domains.
Build core AI platform infrastructure at Harvey including model routing, context engineering, agent infrastructure, and shared evaluation frameworks that power all agentic legal AI products. Requires 8+ years backend experience with 1+ year AI/ML focus and track record of technical leadership.
Builds QA systems, tooling, and workflows to audit and validate large-scale RL training data from suppliers. Partners with vendors to improve data quality using Python, Docker, and AI/ML techniques for frontier AI infrastructure.
Leads AI platform strategy and infrastructure for deploying AI features across Betterment's products, partnering with product, compliance, and engineering teams. Requires deep LLM expertise, fullstack experience with React/GraphQL/server languages, and production AI operations in a regulated environment.
Staff AI Engineer shapes technical direction of AI agents for audit workflows, owning architecture, LLM infrastructure, and reliability systems. Requires 8+ years experience with deep expertise in TypeScript/Python, RAG, and enterprise-scale LLM products.
Builds and scales ML compute platform on Kubernetes with Argo Workflows and Ray for distributed training, orchestration, and resource governance. Optimizes performance, debugs issues, and integrates tooling for ML teams at scale. Requires deep Kubernetes, systems, and programming expertise.
Build and own high-performance inference runtime for Voice AI models including STT, TTS, and voice agents. Design real-time systems with low tail latency, collaborate cross-team, and optimize model serving for production workloads. Requires CS degree and real-time systems experience.
Designs and optimizes embedded computer vision pipelines for real-time detection, tracking, and 3D localization of fast-moving drone threats in autonomous counter-UAS systems. Requires expertise in C++/Python, OpenCV, deep learning frameworks, sensor fusion, and edge deployment on NVIDIA Jetson.
Develop machine learning systems that generate scientifically accurate, editable visuals from biological research and support natural-language figure editing. The role requires deep ML expertise and a passion for solving novel problems, with scientific or research experience as a plus.
Builds and optimizes RL training infrastructure, removes bottlenecks in the RL stack, and partners with researchers to accelerate model development at scale. Requires strong software engineering, ML infra experience, and comfort across the stack.
Research-minded Data Scientist building experimentation platforms, ML analytical systems, and causal inference models for data-driven decisions at Figma. Requires PhD in quantitative field, SQL/Python fluency, and strong stats/ML foundation.
Designs and implements infrastructure for training next-generation LLM architectures focused on Mixture-of-Experts and long-context models. Requires 4+ years building ML pipelines with PyTorch/JAX/TF, deep learning expertise, and strong software engineering skills.
Combines clinical expertise (MD/PA/DO) with AI to develop and refine tools for clinical documentation, decision support, and risk adjustment. Defines quality standards, optimizes prompts, evaluates outputs, and collaborates with engineering teams to enhance clinician efficiency.
Develops generative and programmatic CAD pipelines for micro-device assemblies, integrating AI-driven workflows and CAD APIs to automate design exploration and optimization. Requires CAD scripting expertise, AI knowledge, and 5+ years experience (PhD) or equivalent.
Develops novel machine learning methods for computer vision, leveraging user data insights to innovate, publish at top conferences, and deploy to production. Requires Masters/PhD in relevant field, publications, Python proficiency, and pragmatic research approach.
Research Engineer building production LLM and ML systems for healthcare workflows. Requires strong ML/NLP research background with publications, production deployment experience, and proficiency in PyTorch/TensorFlow/JAX.
Leads ML engineering for 3D scene reconstruction platform, driving quality improvements in production pipelines, managing team of engineers and scientists, and coordinating cross-functionally. Requires 7+ years ML experience including 5+ in management, heavy computer vision expertise, and master's degree.
Builds distributed ML infrastructure including GPU training, end-to-end pipelines, and deployment platforms. Requires 3+ years experience in production ML systems, strong software engineering, and familiarity with open-source tools.
Conducts research on reinforcement learning for self-driving cars and robotics, develops large-scale RL training infrastructure, and deploys algorithms to production systems. Requires hands-on RL experience, PyTorch/Python proficiency, and strong research skills.
Conducts cutting-edge research in reinforcement learning, self-play RL, VLA post-training, and closed-loop RL for autonomous driving and robotics. Requires strong research record with publications, MSc/PhD in ML/CV, and expertise in Python, PyTorch, computer vision, and robotics.
Develops and deploys generative ML techniques for production-grade sensor simulation (Lidar, Radar, Cameras) in autonomous systems. Collaborates with research, rendering, and physics teams; requires 5+ years ML experience, Bachelor's in CS, and expertise in large models and 3D geometry.
Conducts research in reinforcement learning and VLA post-training for robotics and autonomous systems, focusing on dexterous manipulation. Publishes at top conferences and deploys algorithms to real-world products. Requires MSc/PhD, strong publications, and expertise in Python, PyTorch, CV, robotics.
Builds ML tools, infrastructure, and manages large datasets for end-to-end autonomy research and productionizing self-driving software. Works with AI research and engineering teams to scale GPU compute, data, and evaluation systems. Requires strong software generalist skills across ML stack.
Designs, builds, and operates large-scale ML infrastructure for AI/RL research, including GPU cluster orchestration, data curation pipelines, and distributed training systems for autonomous driving and robotics.
Research Engineer focusing on 3D vision, Gaussian splatting, foundation models, and generative techniques for self-driving applications. Conducts cutting-edge research, publishes at top conferences, and deploys algorithms to production autonomy systems. Requires hands-on experience in 3D/ML for AV/robotics, Python/PyTorch proficiency.
Develops ML-first behavior prediction modules to forecast road user motions and interactions for autonomous systems. Requires 3+ years experience with deep learning end-to-end cycles, C++/Python fluency, and collaboration with perception/planning teams.
Defines and manages perception subsystem requirements for autonomous trucking, evaluates safety and performance, coordinates with SW/HW teams, and develops V&V pipelines. Requires 3+ years automotive systems engineering experience with perception and sensor expertise.
Develops cutting-edge robot learning technologies including RL training in simulations, hardware setup, data processing, and end-to-end autonomy algorithms for real-world robotic systems. Requires hands-on experience in multi-modal robot learning, reinforcement learning, or related fields, plus Python, PyTorch, computer vision, and robotics expertise.
Builds and operates scalable data pipelines and systems for post-training workflows, model evaluations, and synthetic data generation. Partners with frontier AI labs and customers, requiring strong backend skills in Python/Go/Rust and ML evaluation expertise.