Latest remote ML Engineering jobs
Job results
Quant Developer designing, implementing, deploying, and owning production vault strategies for onchain finance at Gauntlet. Full ownership of quantitative systems from research through live operation, risk management, and on-call monitoring. Requires production quant systems experience, strong Python, and applied statistics/optimization skills.
Build and operate backend services for AI and model inference on Confluent's real-time streaming data platform. Own end-to-end feature delivery across model lifecycle, inference routing, and agent execution with strong distributed systems expertise.
Founding engineer building an AI-native platform to automate Coinbase's finance workflows (period-close, reconciliation, regulatory filings). Architect governed LLM agents with SOX-compliant controls, integrate with ERP systems, and set technical direction for FP&A/Treasury as an embedded engineer.
Lead a team of ML engineers building and scaling production ML systems for Fetch's Ad Platform, including ad ranking, targeting, bidding, and optimization. Requires 8+ years technical experience (2+ managing teams), strong ML lifecycle and production systems expertise, and cross-functional partnership skills.
Build and manage scalable MLOps infrastructure, automated ML pipelines, model serving, and monitoring for a Large Tabular Model (LTM) at an enterprise AI company. Requires 5+ years MLOps/DevOps experience with Kubernetes, PyTorch/TensorFlow, and cloud infrastructure.
Senior Manager overseeing ML data operations and labeling quality at Coalition. Define guidelines and metrics, drive Label Studio platform requirements, manage vendors, and partner with ML teams to ensure high-quality labeled datasets for models.
Principal Software Engineer building a new Identity Graph platform for real-time identity resolution, fraud, and risk on the Stytch team at Twilio. Own architecture, high-scale distributed systems, complex data pipelines, and synchronous read models while mentoring the team.
Intern on the engineering team owning and shipping a real scoped project end-to-end (design to production code) in voice AI, ASR, TTS or related systems. Requires self-motivated builder with first-principles reasoning, AI-first mindset, and ability to quickly learn new languages/codebases.
Build and ship production-grade AI-powered features and agentic workflows for analytics and large-dataset use cases. The role requires strong software engineering experience, practical GenAI and LLM expertise, cloud-native exposure, and a rapid experimentation mindset.
Principal Software Engineer driving technical vision for Upstart's Core Pricing platform. Lead design of large-scale ML-powered systems to optimize loan pricing, borrower-lender matching, and marketplace efficiency for personal and unsecured loans.
Build and productionize internal agentic workflows and tooling on OpenRouter to automate support and go-to-market operations. Requires build-over-buy conviction, reliability focus, domain knowledge in support/GTM, backend systems expertise, security mindset, and quantitative evals.
Build and productionize cutting-edge ML models for Airbnb's query intelligence, including autocomplete, query tagging, expansion, intent modeling, and LLM-powered natural language search to understand guest intent.
Senior Staff Software Engineer building an AI agent platform and automated workflows to transform Coinbase's Legal organization. Architect production-grade LLM and multi-agent systems that replace manual legal processes such as agreement redlining and governance.
Machine Learning Engineer building statistical models, optimization systems, and experiments for mobile ad tech economics on the Revenue Engine team. Requires PhD in CS/ML/Economics and industry experience applying ML or economics at scale.
Build and own evaluation systems, metrics, pipelines, and datasets to rigorously measure the quality of Firecrawl's LLM-ready web data outputs at massive scale, driving model and product improvements.
Senior Application Engineer building and deploying AI Agents with Salesforce Agentforce, intelligent automations, and custom Salesforce applications. Requires 5+ years experience, specific Salesforce certifications (Platform Developer I/II, Advanced Admin, Agentforce Specialist), strong coding and DevOps skills, and mentoring ability.
Build and own the core AI platform powering an AI-native trading copilot. Develop high-performance Rust backend for streaming, tool execution, and safe trading actions; design robust APIs with observability and security. Requires 8+ years systems programming experience.
Own reliability and quality for an AI copilot in a trading platform. Design evaluation systems, benchmarks, quality gates, model improvement loops, and AI monitoring for correctness, safety, and performance in market analysis and trading workflows. Requires 8+ years production software experience and strong ML eval expertise.
Build high-performance Rust backend and core AI platform powering an AI-native trading copilot with streaming, low-latency tools, safe trading actions, and robust APIs. Requires 8+ years experience in systems programming, production services, and architecture.
Senior engineer responsible for building production-grade Python connectors and AI/ML integrations that enable ClickHouse in RAG, feature-store, and LLM application workflows. Requires 7+ years of software development experience and hands-on Data Scientist or ML Engineer experience.
Senior engineer owning Python connectors and AI/ML integrations that connect ClickHouse to RAG, vector search, feature-store, and LLM application workflows. Requires 7+ years of software development experience and hands-on Data Scientist or ML Engineer experience.
Own the research-to-production pipeline at Deepgram, turning experimental speech ML models into reliable, scalable production services. Partner with researchers on robust workflows, automated release gates, inference optimization, and feedback loops across hybrid GPU infrastructure.
Backend Engineer building and optimizing Deepgram's core inference services for speech processing, including networking, audio transcoding, latency/memory optimization, and distributed compute orchestration. Requires 3+ years experience with Rust (or C/C++) and Python.
Build and optimize infrastructure for frontier-scale reinforcement learning and distributed model training, including kernels, runtimes, parallelism, and asynchronous rollouts. The role requires strong AI/ML systems experience, PyTorch expertise, and GPU performance optimization skills.
Conducts frontier research and builds scalable synthetic-data and distributed reinforcement-learning infrastructure for large AI models. Requires strong AI/ML engineering experience, distributed inference expertise with tools such as vLLM or SGLang, and MLOps knowledge.
Research Engineer building and optimizing distributed infrastructure for frontier-scale model training and reinforcement learning. The role requires strong AI systems experience, PyTorch and distributed-training expertise, GPU performance optimization, and familiarity with parallelism and large-scale clusters.
Senior technical leader building, productionizing and operating large-scale ML models and Agentic AI systems to fight fraud, ensure safety and build trust across Airbnb's platform. Requires 12+ years applied ML experience and deep expertise in LLMs/GenAI.
Build and scale production multi-agent systems for customer support, integrating LLMs with internal tools and APIs. The role requires 3+ years of production ML/AI experience, strong Python and microservices expertise, and familiarity with orchestration, RAG, evaluation, and vector databases.
Build and scale AI-powered workflow automation, including prompt-based actions, custom agents, RAG systems, knowledge bases, and advanced data features. The role requires strong JavaScript and Node.js expertise, AI platform experience, and at least four years of software development experience.
Leads hands-on architecture and implementation for an AI-powered React application builder, including LLM pipelines, distributed backend services, generation systems, and deployment infrastructure. Requires 6+ years of backend engineering experience, strong Go and Node/NestJS skills, and expertise in generative AI, reliability, and scalable systems.
Lead and mentor an MLOps team to build scalable ML infrastructure, automated pipelines, and low-latency model serving for large tabular models. Requires 7+ years MLOps experience including 3+ years leading teams, deep expertise in Kubernetes, model serving frameworks, and observability tools.
Senior technical IC owning the Modeling to ML Serving to API architecture for Airbnb's Host Pricing platform. Lead unified serving stack, backfill/evaluation infrastructure, and domain contracts between ML modeling and serving teams.
Senior Engineer building production-grade applied AI systems, LLM-powered tools, retrieval/memory patterns, and human-in-the-loop workflows to create organizational intelligence for home care operations. Requires Python experience, production system design skills, and cross-functional collaboration.
Build and evolve auction, bidding, and budgeting ML systems that power Reddit Ads. Design optimization algorithms balancing advertiser performance, user experience, and marketplace efficiency.
Hybrid ML/SRE role owning reliability, security, and safety of a large fleet of generative media model APIs (image, video, audio). Build observability for ML-specific failures, harden deployments, operationalize safety systems, lead incident response, and improve GPU fleet efficiency.
Build and deploy cutting-edge ML and Generative AI systems to transform Airbnb's customer support experience, focusing on LLM fine-tuning, RAG, and intelligent service automation.
Build and lead production-grade machine learning services for Coinbase’s conversational AI ecosystem, coordinating vendor and internal LLM systems through a scalable orchestration layer. The role requires 5+ years of ML and software engineering experience, strong Python skills, and expertise in modern generative AI architectures.
Build and improve large-scale ML models for personalization and recommendation across Pinterest surfaces including Shopping. Requires 5+ years applying ML methods and experience with big data pipelines.
Senior Research Engineer building data synthesis, analysis, and management tooling for AI safety model training and evaluation. Requires strong software engineering, statistics, and ML framework expertise.
Senior Engineer building multi-agent AI systems, LLM integrations, and backend automation services that power Marketing Operations. Owns technical direction for agentic infrastructure connecting models to business systems.
Builds and leads production-grade AI-powered conversational systems, including orchestration services connecting LLM frameworks, vendor AI, internal agents, and human workflows. Requires at least five years of machine learning and software engineering experience, strong Python skills, and expertise in modern generative AI architectures.
Senior IC building and maintaining ML underwriting and credit decisioning models for Cash App Borrow and Afterpay. Owns full modeling lifecycle including experimentation, calibration, deployment, and monitoring.
Founding ML Engineer building production ML systems for governance, security, and agentic platform capabilities at Docker. Requires 5+ years applied ML experience shipping systems and 4+ years backend/infra engineering.
Staff ML Engineer leading end-to-end identity verification ML systems including document authenticity, face matching, liveness detection, GNN-based identity graphs, and behavioral risk models. Requires 8+ years production ML experience and domain expertise in biometrics or fraud detection.
Founding Staff ML Engineer building production ML systems for governance, security, and agentic platform capabilities at Docker. Owns architecture, data pipelines, evaluation, and model lifecycle while mentoring the growing team.
8-month modelling residency embedded in production and research projects, focusing on adaptive algorithms, real-time learning, model efficiency, and cross-stack optimization. Requires Python, deep learning frameworks, and ML optimization experience.
Build and fine-tune vision-centric VLMs and generative models using Pinterest's visual-text datasets. Requires 2+ years industry computer vision experience and an M.S. or Ph.D.
Build and operate production multi-agent AI systems that turn longitudinal health data into actionable clinical recommendations. Requires 6+ years building production ML or backend systems plus 1+ years building agentic AI.
Build and productionize large-scale recommendation systems, NLP/embedding models, and agentic AI workflows. Requires 6+ years ML/NLP experience and expertise in PyTorch/TensorFlow, RAG, and vector search.
Staff-level engineer building LLM/ML systems for clinical documentation review, risk detection in healthcare claims, and provider-patient matching at a mental healthcare platform.