Latest ML Engineering jobs
Job results
Senior forward-deployed engineer builds and deploys AI-powered workflows using LLMs and ML frameworks for client operational problems. Requires 7+ years experience in Python, data/ML systems, deployment infrastructure, and client-facing skills; hybrid in NYC area.
Curates, builds, and scales AI-powered drug discovery tools including structure prediction, protein design, and docking models. Collaborates with customers to troubleshoot and optimize biological AI workflows using Python, PyTorch, and cloud infrastructure.
Develop and scale recommendation algorithms, ranking, and search systems using AI stacks and data infrastructure for a platform with 600M users. Requires expertise in recommender systems, deep learning frameworks like JAX/PyTorch, and tools like Kafka/Spark.
Develop and integrate Grok AI models into the Ads Platform to power bidding, auction, ranking, prediction, and optimization systems at massive scale. Requires 3+ years experience with large-scale AI-powered advertising products.
Develops multimodal mid-training data pipelines for omni models handling text, image, video, and audio. Requires ML expertise, scaling laws knowledge, experiment design, and strong engineering in large-scale data frameworks like Spark and Ray.
Leads ML team building scalable systems for personalization, ranking, search, and ads. Owns end-to-end architecture from training to serving at consumer scale, requiring 8+ years ML experience and strong systems design skills.
Senior Forward-Deployed Engineer embeds with federal customers to deploy Voice AI solutions, owning full technical lifecycle from PoC to production. Requires 5+ years experience, TS/SCI clearance, full-stack skills in Python/JS/Rust, and federal deployment expertise.
Builds ML-based systems integrated into Radar's geofencing, maps, and fraud products across backend, data infra, and mobile SDKs. Drives end-to-end features leveraging ML for production-scale impact, collaborates with customers.
Build and optimize scalable machine learning systems for LLM pretraining, fine-tuning, alignment, and evaluation. The role requires 4+ years of ML systems experience, strong Python and PyTorch skills, and familiarity with transformers; experience with LLMs, reinforcement learning, and distributed training is preferred.
Builds and deploys machine learning models to address business challenges. Requires 3+ years experience, Python proficiency, TensorFlow/PyTorch, and production ML deployment skills.
Leads architecture and development of agentic AI systems for personalized brain health therapies, including conversational therapy and education platforms. Requires 8+ years software engineering with 2+ years AI/ML focus, Python/TypeScript expertise, and distributed systems experience.
Develops and optimizes AI inference engine for image/video enhancement, focusing on performance, GPU/CPU optimization, model deployment, and hardware partnerships. Requires C/C++ expertise, 1+ years experience in performance optimization and image processing.
Designs, fine-tunes, and deploys image generation models for photorealistic AI bots, optimizing for consistency, latency, and quality. Requires 5+ years software engineering, 2+ years production ML, and expertise in diffusion models like Stable Diffusion and PyTorch.
Builds and scales ML models/systems for Generative AI content generation (text, image, audio, video). Owns roadmaps, processes for training/fine-tuning at scale, and collaborates cross-functionally. Requires 10+ YOE, Master's/PhD, Python/C++ proficiency.
Build and deploy AI/ML systems including LLM-powered document extraction, agentic workflows, sales forecasting, anomaly detection, and promotion optimization for CPG brands. Requires 2+ years applied AI/ML and 6+ years production software engineering experience.
Builds scalable machine learning systems for real-time fraud detection using unsupervised/supervised ML, big data tools like Spark and Kafka, and streaming technologies. Requires 1-5 years experience in Java, Python, Shell, and big data technologies.
Designs, deploys, and optimizes AI agents to automate construction permitting workflows. Builds backend/frontend services, APIs, data pipelines, and evaluation systems. Requires 3+ years in software/ML engineering with production AI experience.
Build and optimize large-scale ML systems for safe, steerable AI, handling infrastructure, experiments, and dev tooling. Requires strong software engineering and interest in ML research.
Builds machine learning systems using unsupervised/supervised algorithms and deep learning to detect fraud in real-time. Designs distributed streaming systems with big data technologies like Spark, Kafka, and Flink. Requires 5+ years in Java, Python, Shell.
Leads engineering initiatives to build scalable ML platform infrastructure, including unified embeddings, feature pipelines, and continuous learning systems. Requires 7+ years in applied ML, Python/ML frameworks expertise, and Master's/PhD in quantitative field.
Builds production-grade AI systems including agent architectures, evaluation pipelines, and workflows for enterprise demand intelligence platform. Owns end-to-end from prototype to deployment, focusing on accuracy, latency, and cost tradeoffs in ambiguous startup environment.
Build and own autonomous agentic systems for customer-facing revenue and marketing workflows, partnering with sales and marketing teams. Requires 4+ years experience in software/ML engineering, full-stack skills in Python/JavaScript, and production systems expertise.
Senior/Staff Software Engineer architects ML data pipelines for autonomy AI at Nuro, transforming massive datasets into training signals. Requires 7+ years experience, C++/Python proficiency, and expertise in end-to-end ML data systems.
Leads cross-functional strategic projects in Generative AI to drive multimillion-dollar revenue, owning data labeling operations and product enhancements. Requires 2+ years experience, strong technical skills in SQL/Python, and entrepreneurial mindset.
Builds AI-powered concierge systems for luxury travel including recommendation flows, RAG for searches, preference engines, and integrations. Requires 3+ years cloud/AI experience with LLMs, vector search, and production systems; hybrid in SF office.
Leads design, customization, and integration of LLMs into biomedical research workflows and NCBI platforms like PubMed. Requires 3+ years hands-on LLM experience with Python, PyTorch, Hugging Face, and RAG systems; biomedical domain preferred.
Builds and launches AI-powered workflow automation products for government agencies, owning full development from needs identification to production using LLMs, RAG, and multi-agent systems. Requires 5+ years experience with generative AI and product engineering in startups.
Builds and deploys production AI systems including LLM-powered features, agents, and RAG for underwriting, pricing, and risk evaluation in insurance. Requires 5+ years backend experience, hands-on AI product shipping, and familiarity with LLMs and vector DBs.
Founding ML Engineer builds ML strategy, production systems, and infrastructure from scratch for AI cybersecurity threat detection using LLMs. Requires 8+ years production ML experience, strong software engineering, and cloud/ML frameworks; leads team growth.
Partners with sales to drive technical sales of AI/ML infrastructure, leading demos, POCs, and solutions for enterprise customers. Requires 2+ years software engineering, AI/ML expertise, and strong communication skills.
Develops and launches machine learning algorithms powering Lyft's recommendation systems and core services, partnering with cross-functional teams to solve diverse problems in transportation and personalization. Requires 5+ years ML experience, Python/Golang proficiency, and advanced ML methodologies.
Senior/Staff Software Engineer builds end-to-end AI products for clients, spending 75% time coding and 25% with stakeholders like CTOs. Requires 8+ years experience shipping products, applying software engineering to AI systems with high ownership.
Leads design and deployment of AI/ML models, prototypes, and ethical frameworks for DoD missions. Requires 8+ years in AI/ML/data science, bachelor's in relevant field, Python/TensorFlow/PyTorch experience, and Public Trust eligibility.
Build and operate scalable ML infrastructure for deploying models to IoT sleep devices. Own end-to-end pipelines, optimize performance, and collaborate cross-functionally. Requires 5+ years in ML ops, Python, AWS, and production ML deployment.
Develops and scales MPI+CUDA PDE solvers and neural operators for electromagnetic simulations on GPU clusters for IC design. Requires PhD in computational physics/math, expertise in numerical methods, C++/CUDA, HPC, and neural operators for physics problems.
Founding Engineer builds full-stack Shopify app, AI video extraction pipelines, and creator tools to create commerce evidence layer from buyer videos. Requires AI-native full-stack experience, strong fundamentals, and Bay Area location.
Build system-level debugging, validation, observability, and anomaly-analysis tooling for Cerebras’s AI hardware and software stack. The role requires strong C++ and Python skills, experience debugging complex hardware/software systems, and familiarity with compilers, runtimes, or high-performance computing.
Build ML systems for perception data in autonomous vehicles, curating diverse datasets, synthetic data frameworks, and tools to identify data gaps. Requires 4+ years software engineering with Python/C++, ML implementation experience, and BS in quantitative field.
Builds and operates scalable production ML systems for real-time personalization and targeting. Requires 6+ years experience with Python, PyTorch/TensorFlow, data pipelines, and cross-functional collaboration.
Builds and optimizes distributed frameworks for LLM training and inference on Scale's RLXF platform. Collaborates with ML teams to accelerate research, requiring expertise in PyTorch, CUDA, transformers, and large-scale distributed systems.
Staff engineer owns large product areas in enterprise GenAI platform, working across backend, frontend, LLMs, and ML models. Solves scalability challenges with 7+ years experience in Python/JS, Kubernetes, and cloud providers.
Build and scale full-stack systems for Scale AI's Generative AI Data Engine, owning contributor platform features that power high-impact datasets for LLMs. Requires 5+ years experience with React, TypeScript, Node.js, and strong product sense in a hybrid SF/NY environment.
Leads development and optimization of distributed frameworks for LLM post-training, training, and inference. Collaborates with ML teams to enable advanced model development and data curation, requiring expertise in large-scale ML systems and tools like PyTorch and CUDA.
Designs, builds, and deploys production-ready AI agents using LLMs, tool use, and reasoning for enterprise problems. Requires 5+ years ML experience, Python proficiency, and Bachelor's in CS/ML/AI.
Develops synthetic data pipelines, production trace agents, and automated agent-building frameworks for enterprise GenAI. Requires 3+ years LLM production experience, top conference publications, and advanced CS degree.
Build and scale enterprise Generative AI platform, owning large product areas across backend, frontend, LLMs, and ML models. Requires 4+ years experience, proficiency in Python/JavaScript/SQL, Kubernetes, and cloud providers.
Software Engineer on Claude Code team builds evaluation systems, tooling, and infrastructure to enhance AI coding capabilities. Collaborates with researchers in fast-paced environment; requires 5+ years experience building complex systems.
Designs and optimizes TPU kernels to address performance issues in ML research, training, and inference systems. Provides feedback on model impacts and solves large-scale systems problems, requiring deep accelerator expertise.
Build ML systems to detect and mitigate AI misuse, including classifiers for anomalous behavior, multi-exchange harm monitoring, and agentic safety evaluations. Requires 4+ years ML experience, Python proficiency, and research-to-deployment skills.
Software Engineer focused on AI reliability engineering, improving robustness of Claude's serving infrastructure across SDK to accelerators. Partners cross-team on SLOs, monitoring, high-availability systems, incident response, and safeguard models.