Latest ML Engineering jobs
Job results
Designs, builds, and deploys production-grade AI agents using LLMs for enterprise workflows. Collaborates with customers to solve business problems, ensures reliability, and mentors teams on agentic architectures.
Builds and deploys end-to-end AI/ML systems, including LLM workflows and rapid prototypes for internal tools and product features in healthcare. Requires 4+ years experience, strong full-stack skills, evaluation/monitoring expertise, and Python proficiency.
Build and maintain scalable ML platform for model experimentation, training, evaluation, inference, and feature store to power underwriting products. Requires 5+ years experience with Python, ML stacks like Databricks/AWS, and MLOps systems.
Designs, builds, and scales agentic AI systems using LLMs for compliance automation, including multi-step reasoning, RAG, and production deployment. Requires 7+ years software engineering with 2+ years ML/AI, Python proficiency, and cross-functional collaboration.
Builds sophisticated AI agent products for investigating financial crimes, leads technical designs and teams. Requires 4-10 years experience in software development, Python, cloud infrastructure, and hybrid SF Bay Area work.
Leads design and development of scalable, real-time API and data infrastructure for AI model evaluations, processing large-scale event streams with low latency. Requires 5+ years in infrastructure or ML systems, expertise in distributed systems, stream processing, and backend architecture.
Builds and optimizes ML backend systems for recommendation, ranking, and search in a fast-growing AI consumer app. Requires 5+ years experience, CS bachelor's, ML frameworks like PyTorch/TensorFlow, cloud infra, and modern typed languages.
Develops ML models using imitation and reinforcement learning for agent behavior prediction and autonomous vehicle planning. Requires PhD +1 year or MSc +5 years experience, expertise in RL, transformers, and production ML pipelines.
Builds personalization and recommendation systems from scratch for AI creative tools, modeling user taste and curating feeds. Designs algorithms for generative models to adapt to individual aesthetics, using ML frameworks like PyTorch/JAX.
Builds and deploys AI primitives and agents to automate workflows and enhance user experiences in investment management platform. Requires AI agent experience, distributed systems knowledge, and product-minded engineering across tech stacks.
Designs, trains, and deploys deep learning models for drone autonomy, focusing on computer vision tasks like optical flow, depth estimation, detection, and path planning using real-world and synthetic data. Requires MS/PhD, hands-on DL experience, and PyTorch/Python/C++ proficiency.
Builds and scales deep learning infrastructure for autonomous drone computer vision workloads, optimizing inference for high throughput/low latency across hardware, and implements MLOps pipelines for model deployment and monitoring. Requires hands-on MLOps, DL/CV expertise, and ML pipeline experience.
Develops real-time deep learning models for drone autonomy, focusing on detection, tracking, segmentation, and optical flow using visual data. Requires hands-on experience with deep neural networks, computer vision, Python/C++, and model optimization for embedded hardware.
Develops and optimizes deep learning inference infrastructure for real-time computer vision workloads on drones, focusing on high-throughput, low-latency performance across hardware platforms. Builds MLOps workflows, GPU kernels, and SDKs; requires strong DL, CV, and ML pipeline expertise.
Builds and improves AI agents for complex accounting tasks, working across product engineering, agent platforms, infrastructure, and data systems at the frontier of applied ML. Requires strong systems thinking, ownership, and excitement for coding agents in a fast-changing environment.
Develops quantitative models for portfolio optimization, fixed income relative value, risk estimation, and AI agents for credit research and portfolio management at a fintech platform. Requires Python coding, quant background, and bachelor's/PhD in math-related field.
Senior Software Engineer building autonomous AI agents for sales workflows (prospecting, outreach, follow-ups). Requires 5+ years experience, production AI agent experience, high agency, and stack familiarity with React, TypeScript, Node.js, Python. Hybrid role with customer collaboration.
Design, train, and integrate ML models for semantic map element detection in autonomous vehicles. Requires 5+ years experience, MS/PhD in CS, expertise in computer vision, deep learning, and PyTorch.
Build and evolve the distributed training framework and tooling powering frontier-scale language models. The role focuses on large-scale ML systems, HPC infrastructure, performance optimization, reliability, and developer tooling across multi-node GPU clusters.
Build and optimize synthetic data and inference pipelines for large language models, combining research and software engineering to improve data quality, throughput, and model performance. The role requires strong Python and data-pipeline experience, familiarity with LLM inference frameworks, and experience with large-scale datasets.
Develops and deploys ML models for parsing unstructured enterprise data like PDFs, focusing on training vision models, experimenting with LLMs, building data pipelines, and integrating into products. Requires 2+ years in production ML, Python proficiency, and computer vision expertise.
Builds and optimizes backend APIs and pipelines for document parsing using LLMs, handling PDFs/spreadsheets at scale. Requires 2+ years experience, exceptional Python, and high agency in production AI systems.
Build and optimize Cerebras’s production GPU prefill and inference stack across APIs, serving runtimes, ROCm, distributed systems, and hardware. The role requires 5+ years of software engineering experience, strong C++ and Python skills, and hands-on experience operating high-performance model-serving systems.
Builds and deploys fine-tuned LLMs and AI agents for real-time voice interactions in consumer lending, ensuring compliance and scalability. Requires 2+ years production ML/AI experience with Python, PyTorch/TensorFlow, and LLM frameworks.
Partners with customers to design, prototype, and scale AI-enhanced coding workflows using OpenAI Codex. Serves as technical expert, leads workshops, builds demos, and influences product direction. Requires 5+ years in technical consulting or solutions engineering.
Build scalable AI platforms and infrastructure for Figma's design tools, including model training, agentic features, and APIs. Requires 5+ years software engineering experience with backend/infrastructure and 3+ years in AI or developer platforms.
Build and productionize ML models for search, RAG, and generative AI features at Figma. Requires 5+ years software engineering with 3+ years in applied ML, Python proficiency, and experience with scalable data pipelines.
Develops machine learning models to predict and optimize organoid growth and differentiation protocols using biological data. Requires Master's/PhD in CS/engineering/math, Python/R proficiency, and ML frameworks like TensorFlow/PyTorch, with biology lab experience.
Designs, develops, and deploys scalable ML systems using LLMs to process clinical data for healthcare applications. Requires 5+ years backend/cloud experience, Python fluency, and familiarity with ML frameworks; works onsite in Boston or NYC.
Designs and implements LLM orchestration frameworks and agent reasoning systems for adaptive mission planning in multi-domain unmanned systems. Optimizes AI models for edge deployment on autonomous vehicles, integrating with ROS autonomy stack for mission-critical operations.
Build and scale agentic AI systems for mission-critical public-interest applications while researching and shipping state-of-the-art models. The role requires strong software engineering, Python and ML framework expertise, LLM experience, distributed GPU training knowledge, and Canadian citizenship with security-clearance eligibility.
Develops high-performance audio inference systems, optimizing latency, throughput, and quality for real-time streaming workloads. Requires expertise in C++, Python, and deep learning models for audio/speech, with collaboration across training and serving teams.
Develops and deploys techniques to enhance LLM inference efficiency, focusing on architecture optimization, decoding algorithms, and GPU acceleration. Requires PhD in ML, expertise in LLM optimization, strong software skills, and top-tier publications.
Engineers on this team optimize LLM inference for lower latency and higher throughput by identifying bottlenecks, developing optimizations across the execution stack, and collaborating with modeling teams. Requires 5+ years high-performance coding in C++/Python and LLM inference experience.
Trains frontier LLMs on semiconductor design/verification data (RTL, netlists, PDKs) for automated chip development. Develops synthetic data generation, model distillation, evals, and scales training across thousands of GPUs.
Post-trains frontier AI models using reinforcement learning to autonomously handle semiconductor design tasks like chip architecture optimization, RTL code generation, simulations, and verification. Collaborates with hardware experts to build RL environments, reward functions, and evaluation frameworks.
Develop and prototype AI-driven solutions for GTM, Finance, and People teams, translating business problems into impactful prototypes using ML/AI and LLMs. Requires 4-8 years as Software Engineer or Data Scientist with production AI familiarity.
Build and maintain scalable backend infrastructure for Fireworks AI's generative AI platform, including LLM CI/CD pipelines, control planes, and model serving systems. Requires 5+ years software engineering experience focused on ML/infrastructure, strong Python/Go skills, and familiarity with PyTorch, Kubernetes, and LLM concepts.
Lead technical development of ML algorithms for next-generation ML Planner. Drive innovations in imitation learning, reinforcement learning, and model scaling while mentoring ML developers.
Build and deploy LLM-powered agents and production AI/NLP pipelines, focusing on RAG, agentic systems, model optimization, and scalable deployment. The role requires 3+ years of machine learning or applied AI experience, strong Python skills, and experience with modern ML frameworks and infrastructure.
Develops agentic AI platforms for triaging, debugging, and resolving production issues using Sentry's error datasets. Requires 5+ years experience, Python/TypeScript proficiency, PyTorch, and expertise in scalable ML deployment.
Develops and optimizes internal distributed ML training framework to boost hardware efficiency and enable researchers to experiment with new AI models. Requires strong Python skills, systems understanding, and passion for performance tuning.
Machine Learning Engineer optimizes ML models for speed and efficiency through low-level CUDA kernel tuning, GPU scheduling, and hardware-aware systems design. Requires 2+ years in ML infrastructure with Python/C++/Rust and distributed frameworks like PyTorch.
Advances small, high-performance language models for retrieval, application, and code generation. Requires 2+ years ML research/production experience, Python/PyTorch/JAX fluency, optimization expertise, and advanced degree.
Builds and productionizes LLM-powered agent workflows, focusing on orchestration, evaluation, reliability, safety controls, and product iteration. Requires strong agentic systems expertise and engineering skills for scalable AI systems.
Build and productionize AI-powered product features for scientific communication, spanning full-stack development, LLM applications, and machine learning infrastructure. The role requires 7+ years of full-stack experience, strong JavaScript and Python skills, and hands-on experience deploying AI systems in production.
AI Enablement Engineers guide enterprise teams in adopting Devin AI software engineer, leading workshops, pair programming on real projects, and scaling enablement programs globally. Requires 3+ years software engineering experience in Python/JS, strong communication, and customer-facing skills.
Builds and optimizes LLM-driven features, agentic workflows, and proprietary AI models for Bubble's visual app development platform. Requires Master's/PhD + 2+ years or 5+ years ML/software experience with transformers, RAG, and AI tools.
Builds production machine learning capabilities for autonomous vehicle perception, prediction, and planning systems. The role requires deep learning expertise, strong C++ or Python skills, and at least three years of production software experience.
Develop and optimize real-time multi-sensor fusion and state estimation algorithms (Kalman/particle filters, IMUs, radar, cameras) for autonomous X-BAT VTOL drone operation in contested environments on the GNC team.