Latest ML Engineering jobs
Job results
Develop and enhance the AI Gateway platform, building unified APIs for AI models with rate limiting, failovers, and integrations. Requires 5+ years experience in JavaScript/TypeScript, backend, and distributed systems.
Develop scalable ML services for data enrichment, managing the full lifecycle from model training with PyTorch/TensorFlow to deploying optimized inference using ONNX/vLLM. Requires 5+ years experience in production ML systems, strong deployment skills, and software engineering proficiency.
Build full stack AI/ML features for finance automation platform, owning greenfield projects with Python/TypeScript/Next.js. Requires 3-5 years full stack exp, 1+ year AI/ML inference, in-person NYC.
Develop backend systems, statistical models, and experiments to optimize Pinterest's ads marketplace, balancing short- and long-term objectives. Requires 10+ years experience, CS/ML degree, and strong engineering/math skills.
Develop and evolve machine learning technology stack for Pinterest's Ads monetization, building personalized recommendation systems using deep learning, LLMs, and big data tools. Requires 2+ years ML experience in recommender systems, ranking, and large-scale systems.
Develops and evolves machine learning technology stack for Pinterest's Ads monetization, building personalized recommendation systems using deep learning and LLMs. Requires 2+ years ML experience, CS degree, and expertise in large-scale systems.
Develop advanced ML models using deep learning to personalize Pinterest experiences across Homefeed, Ads, Search, and more. Requires 4+ years in ML, experience with large-scale systems like Spark/Hadoop, and a degree in CS/ML.
Build and improve machine learning models for Pinterest's recommendation systems across Homefeed, Ads, Search, and more. Requires 2+ years experience in ML methods like personalization and recommender systems, plus hands-on work with large-scale data pipelines.
Senior ML engineer leads development of GenAI models and pipelines for Airbnb's customer support, productionizing at scale. Requires PhD, 10+ years ML experience including 2+ in GenAI, and expertise in NLP, deep learning, and agile AI practices.
New grad Software Engineer for AI Platform teams (Data Platform, Onboard Systems, Technical Infrastructure) at Nuro's self-driving tech company. Build scalable systems for data management, onboard autonomy, and infrastructure using C++, Python, and ML expertise.
Research Engineer building evaluations, search systems, and training agentic ML models for tool integration in AI agents. Requires strong research execution, rapid prototyping, and collaboration to productionize ideas.
Designs and scales distributed data infrastructure for large-scale multimodal AI training and evaluation. Collaborates with researchers to build reliable, high-performance systems in a fast-paced environment.
Builds production AI agents and LLM pipelines for marketing workflows, integrating with data warehouses. Requires strong backend architecture skills, product thinking, and creativity with LLMs; senior role emphasizing impact over years of experience.
Build and scale AI agents automating healthcare workflows. Lead complex feature development using JavaScript/TypeScript/Node.js, LLMs, and DevOps tools. Requires 4+ years experience and CS degree.
Design and deploy low-speed motion planning and control modules for autonomous vehicles, including vehicle-dynamics characterization and optimal-control solutions. The role requires 3+ years of production software experience, hands-on vehicle testing, strong control and numerical-analysis knowledge, and high-performance C++.
Leads development of agentic AI platform to autonomously resolve healthcare insurance claim denials using multi-agent workflows, RAG systems, and LLMs. Requires 5+ years in production ML engineering with Python and frameworks like PyTorch.
Leads design, development, and deployment of AI agents using language models and ML frameworks. Drives scalable, safe AI systems while mentoring teams and collaborating cross-functionally. Requires strong Python skills and AI leadership experience.
Build and improve ML systems and data pipelines infrastructure to support modeling engineers. Requires 2+ years experience, bachelor's in CS/math/sciences, strong coding in Python/Go/Java/C++, and production ML infra expertise.
Optimizes and builds production inference systems for large language models at scale using PyTorch and high-performance tooling. Requires 3+ years experience in production code, OS concepts, and AI inference systems.
Develops and deploys state-of-the-art GenAI models and systems for Databricks products like Assistant and Genie. Requires 2-8 years ML engineering experience, proficiency in Python/PyTorch/TensorFlow, and expertise in LLMs.
Designs and deploys reinforcement and imitation learning algorithms for robotic manipulation tasks in dynamic environments. Requires MS/PhD, deep RL/IL expertise, PyTorch proficiency, and real-world ML deployment experience in a fast-paced startup.
Develops and deploys ML-based perception algorithms for robot object detection, pose estimation, tracking, and scene understanding. Integrates sensor fusion from cameras, LiDAR, IMUs; requires MS/PhD in ML/CS/robotics, expertise in computer vision, PyTorch/TensorFlow, and real-world robotics hardware.
Designs, builds, and owns core cognitive architectures for enterprise AI agent platform, blending AI research with production systems. Requires 7+ years engineering with 2+ years AI/ML production experience and deep Python backend expertise.
Build AI agent systems for legal workflows, optimizing performance via prompt engineering, model selection, tools, and evals. Requires 3+ years experience, Python proficiency, and LLM/agent framework expertise for mid/senior/staff levels.
Designs, implements, and optimizes high-performance GPU kernels for GenAI inference stack. Leads performance improvements, mentors engineers, and collaborates with ML and systems teams. Requires deep kernel programming and GPU architecture expertise.
Leads architecture, development, and optimization of GenAI inference engine for high-throughput, low-latency LLM serving. Requires 6+ years in performance-critical systems, deep ML inference expertise, CUDA/GPU programming, and distributed systems.
Designs and builds scalable, low-latency model serving infrastructure for AI/ML models across CPU/GPU workloads. Requires 10+ years in large-scale distributed systems and deep expertise in inference systems, architecture, and cross-team collaboration.
Designs and builds scalable, low-latency systems for serving frontier AI models on GPUs. Requires 10+ years in large-scale distributed systems, strong system design skills, and leadership in operational excellence; no prior AI experience needed.
Designs and builds scalable infrastructure for high-throughput, low-latency AI/ML model serving on CPU/GPU. Requires 5+ years in distributed systems, inference expertise, and strong system design skills.
Builds high-performance infrastructure for brain reverse-engineering via large-scale AI models and neuroscience-informed simulations. Architects systems, optimizes RL environments, and scales research prototypes to production.
Leads development of Airbnb's conversational AI Automation Platform and agent provisioning systems. Drives backend optimization, prompt engineering, and LLM integrations with 9+ years experience in scalable architectures.
Build evaluation infrastructure for AI systems at Sentry, designing datasets, benchmarks, and test harnesses to measure accuracy and reliability of debugging agents. Requires 5+ years experience, Python/TypeScript proficiency, and AI/ML background.
Builds large-scale ML infrastructure including GPU clusters, training frameworks, and workload schedulers. Requires strong systems engineering in Python, Rust, Golang, with Kubernetes and distributed systems experience.
Build and operate production AI systems including LLM-powered agents, model orchestration, RAG, and backend APIs for tax automation workflows. Requires 7+ years experience with 2+ in LLMs, backend fluency, and startup intensity.
Designs and trains embedding and retrieval models for AI agents to access web data at hyperscale, balancing research innovation with production efficiency for sub-second latency and fresh indexes.
Develops and optimizes inference engines for multimodal AI models, integrating new architectures, building scheduling systems, and managing large-scale GPU deployments. Requires strong Python, model serving frameworks like PyTorch/vLLM, and Kubernetes expertise.
Lead GenAI/ML initiatives for Datadog's APM team, designing and deploying models for agentic workflows, automated investigations, and performance optimization. Requires 10+ years experience with 4+ years leading cross-team projects.
Lead development of LLM observability features at Datadog, building tools for monitoring, tracing, and evaluating AI systems in production. Requires expertise in distributed systems, GenAI applications, and model internals.
Builds and deploys ML models for sleep personalization, readiness forecasting, and behavior prediction using foundation models and multimodal data. Requires 2+ years production ML experience with PyTorch/TensorFlow, personalization systems, and strong Python engineering.
Lead the design, build, and operation of multi-agent AI systems from prototypes to production, owning architecture, reliability, and performance for consumer AI products. Requires strong production experience with LLM-powered systems, agentic workflows, and backend engineering.
Develop and enhance TypeScript-based AI SDK for building AI-native products and agents. Collaborate with distributed engineering team; requires 5+ years experience in JavaScript/TypeScript and open source contributions.
Search machine learning intern contributing to retrieval, ranking, classification, and RAG systems that improve search quality. Requires machine-learning foundations, Python project experience, and enrollment at a Serbian university or equivalent demonstrated software-engineering ability.
Build and operate high-performance inference infrastructure for large language models, deploying optimized NLP models to production with low latency and high throughput using Kubernetes and cloud platforms. Requires 5+ years experience in scalable distributed systems and GPU workloads.
Build and ship full-stack AI projects including AI agents, RAG, structured extraction, and LLM infrastructure. Requires proficiency in full-stack development, backend systems, cloud infrastructure, and production LLM experience.
Builds scalable ML systems and end-to-end pipelines for fraud detection, anomaly detection, and real-time decisioning in payments. Requires 5+ years ML production experience, including 2+ in fraud/risk, Python proficiency, and ML frameworks like PyTorch/TensorFlow.
Leads AI/ML organization by setting technical vision, building and mentoring data scientists/ML engineers, architecting scalable ML infrastructure, and hands-on prototyping models for healthcare applications. Requires PhD, 8+ years shipping ML products, and 3+ years leading ML teams.
Designs and builds scalable backend systems and applies machine learning to address safety, integrity, and Generative AI risks. Requires 8+ years backend experience, ML expertise, and distributed systems knowledge.
Designs, builds, and deploys custom AI agents for enterprise customers, providing hands-on technical support, troubleshooting, and acting as technical account manager. Requires 3+ years customer-facing experience and expertise in Python, AWS, SQL, and AI agent frameworks like LangChain and CrewAI.
Designs and builds AI agents, workflows, and evaluation systems for enterprise audit platforms. Requires experience shipping production software with TypeScript, React, Python, Postgres, LLMs, RAG, and agent orchestration.
Builds performance benchmarking, diagnostic, and optimization tools for LLM inference on GPU clusters. Early-career role requiring Python proficiency, systems curiosity, and interest in AI hardware—no prior experience needed.