Latest Data & AI jobs
Job results
Lead fleet data analysis for V-BAT UAV systems. Own pipelines, metrics, and studies on sensor, navigation, and aircraft performance using large real-world and simulation datasets to drive engineering insights.
Build and maintain large-scale distributed inference systems serving Claude to millions of users. Design intelligent routing, autoscaling, and deployment pipelines across diverse AI accelerators while maximizing compute efficiency for production and research workloads. Requires significant distributed systems experience.
Research-oriented Machine Learning Scientist developing multimodal ML models (NLP, Speech, Computer Vision) for lifelike voice agents. Requires published papers in large-scale deep learning and familiarity with SOTA AI.
Characterize, analyze, and optimize performance of state-of-the-art AI models on Cerebras' wafer-scale hardware. Build performance models, optimize kernels and compilers, debug runtime behavior, and develop visualization tools to influence next-gen AI architecture.
Senior ML Engineer building observability, evaluation frameworks, and improvement loops for production agentic AI systems. Requires 5+ years production ML/LLM experience, strong grounding in agent design or evaluation, and hands-on work taking systems from prototype to scale.
Lead Analyst owns growth measurement, marketplace models, experimentation framework, and analytics for a restaurant loyalty/payments platform. Requires 8+ years in growth analytics or quantitative strategy, expert SQL/Python/R, deep experimentation experience, and fluency in marketplace unit economics.
Build and optimize the RL training framework and infrastructure for large-scale workloads at SpaceXAI, from ablations to production runs. Requires experience with distributed systems and proficiency in Python, JAX, Rust, or C++.
Lead full-funnel marketing analytics and a small AI-augmented team for a consumer fintech platform. Own test-and-learn programs, evolve measurement frameworks (MMM, MTA, incrementality) with ML partners, and drive insights to optimize CAC, LTV, and growth.
Build and operate production data pipelines, observability tools, and planning systems to maximize utilization, efficiency, and attribution of Anthropic's large-scale multi-cloud accelerator and CPU fleet. Requires strong Python/SQL, cloud operations, and Kubernetes experience in a high-ambiguity environment.
Research Engineer on OpenAI's Privacy team designing and prototyping privacy-preserving ML algorithms like differential privacy and federated learning at scale. Requires hands-on PETs experience, fluency in PyTorch/JAX, and a track record implementing or publishing novel privacy work.
Conduct original research on LLM evaluation, routing optimization, and model behavior using billions of real-world generations. Design novel benchmarks, run large-scale empirical studies, and develop statistical foundations for intelligent routing. Requires MS/PhD, publication track record, deep stats/ML expertise, and Python/SQL skills.
Forward Deployed Data Engineer building hybrid data pipelines and semantic layers for Hilbert's AI Growth Engine. Implements warehouse-native or managed ClickHouse integrations, partners with AI agents for accelerated onboarding, and ensures reasoning consistency across customer environments.
Research Engineer building self-improving AI agent systems at Console. Develop eval/optimization loops, fine-tune specialist models, and improve agent reasoning over enterprise context using production data to drive measurable gains in quality, latency, and reliability.
Build and optimize scalable AI infrastructure for real-time inference, evaluation, and continuous improvement of LLMs, LVMs, computer vision, and multimodal models on large-scale video data. Requires 4+ years production ML systems experience, strong Python skills, and expertise in inference optimization and model serving.
Senior Platform Engineer designing, building, and scaling database infrastructure for production and AI systems, including relational, analytical, and vector stores. Requires 7+ years experience with distributed systems, high-availability databases, and supporting AI workloads like vector search and RAG.
Build and own validation pipelines, CI/CD infrastructure, and platform integrations to launch frontier models and inference features reliably across AWS, GCP, and Azure. Requires strong large-scale distributed systems experience and track record improving release velocity.
Quantitative Intelligence Analyst discovering novel risks in human-AI systems using quantitative models, data mining, and statistical analysis to develop early-warning signals for trust and safety threats. Requires 3+ years in quantitative intelligence, risk research or trust & safety, plus Python/SQL skills.
Senior Data Engineer building scalable data pipelines, services, and AI-ready data layers in Go/Scala/ClickHouse to power internal products, analytics, and agentic AI for go-to-market, engineering, and product teams. Requires 5+ years experience in production data systems, strong programming, SQL, and databases.
Build and scale the shared AI platform foundations at Notion, enabling fast and safe shipping of AI products. Requires experience with LLM/ML platforms, strong ownership, and comfort across backend, infrastructure, and product code.
Operations Analytics Manager building and owning the end-to-end analytics stack for demand forecasting, S&OP, supply chain optimization, and costed BOM analytics. Requires 5+ years in data science or operations analytics, Python/SQL proficiency, forecasting model experience, and cross-functional partnership with leadership.
Help build and improve HappyRobot’s AI agent products by defining metrics, designing and interpreting A/B tests and experiments, analyzing conversations/workflows, and turning data into insights that guide Product and ML roadmaps. Requires 4+ years in data science or product analytics, strong SQL/Python, stats, and experimental design experience.
Senior Data Engineer responsible for building scalable ingestion pipelines, normalizing, and maintaining large-scale financial and alternative datasets from global vendors to support quantitative research and alpha generation. Requires 5+ years data engineering experience in finance/quant environments, strong Python/SQL/Linux skills, and deep knowledge of market/tick/reference data across asset classes.
Senior Staff Software Engineer building an AI agent platform and automated workflows to transform Coinbase's Legal organization. Architect production-grade LLM and multi-agent systems that replace manual legal processes such as agreement redlining and governance.
Research Engineer advancing RL for silicon chip design at Anthropic. Design RL environments for RTL generation, verification, and physical optimization; requires deep ASIC/FPGA expertise from spec to tapeout.
Machine Learning Engineer building statistical models, optimization systems, and experiments for mobile ad tech economics on the Revenue Engine team. Requires PhD in CS/ML/Economics and industry experience applying ML or economics at scale.
Data scientist owning analytics and experimentation for new user activation and habit formation in Suno's early-life journey. Requires 5-7+ years in consumer data science/product analytics, advanced SQL/Python/statistical modeling, and deep experimentation experience.
Build synthetic data pipelines and tasks to train frontier AI agents. Requires Python, Docker, Linux, and experience creating realistic, scalable synthetic training data for models and evals.
Build high-quality, domain-specific benchmarks and infrastructure to rigorously evaluate frontier AI agents on realistic workflows. Requires strong Python/Docker/Linux skills, experience with evals or benchmarks, and a deep understanding of what makes a benchmark reliable and useful.
Build QC automation systems for RL training data and agent evals at HUDHUD. Design quality standards, validation pipelines, experiments and metrics without heavy LLM reliance; partner with vendors to debug and improve data generation. Requires Python, Docker, Linux and experience building scalable QA/QC systems end-to-end.
Applied Research Engineer owning technical deployment requests, troubleshooting, building tools/pipelines, and resolving ambiguous problems for frontier AI labs and data vendors at HUDHUD. Requires strong Python/Docker/Linux skills, eval/benchmark experience, independent problem-solving in fast-paced ambiguous settings.
Senior Data Analyst partnering with Marketing on performance measurement, attribution, funnel analysis, forecasting, and lifecycle metrics in B2B SaaS. Build data foundations for AI/automation, translate business questions into analytics, and drive data-informed decisions using advanced SQL, cloud tools, and martech platforms.
Software Engineer building scalable data infrastructure, cataloging, versioning, and lineage tools to support ML research and production workflows at an AI-driven hedge fund. Requires 3+ years experience, strong software design skills, and expertise in a modern language like Python or Java.
Senior analytical partner to the Marketing organization at 1Password. Define marketing performance measurement, influence leadership investment decisions, shape data foundations for AI/automation, own attribution/funnel/forecasting/lifecycle analysis in B2B SaaS. Requires 7+ years marketing analytics experience and deep domain expertise.
Staff Machine Learning Engineer developing state-of-the-art visual encoders and multimodal models at Pinterest Labs. Prototype visual reasoning tools, train billion-scale models on rich visual-text data, ship to production for recommender systems and VLMs, publish research, and mentor juniors. Requires strong CV/ML background, publications, and PhD or equivalent.
Lead a team of data scientists focused on marketing analytics for retention and acquisition. Oversee production ML models, causal inference projects, and LLM integration to drive marketing decisions and business outcomes. Requires 7+ years data science experience including 3+ years managing teams.
Lead AI evaluation for Figma's AI-powered products. Define quality metrics, build human + automated eval frameworks (rubrics, golden datasets, LLM-as-judge), manage a small team, and deliver decision-ready insights to Product, Design, and Engineering stakeholders.
Build and own ML models, fine-tuning, evaluation harnesses, and routing for Kepler's AI agent harness in finance. Requires 5+ years production software experience and shipped ML systems focused on correctness, evals, and real-world reliability.
Build and deploy machine-learning perception systems for L4 autonomous trucks, spanning multimodal 3D detection, sensor fusion, tracking, localization, and safety validation. Requires 5+ years of perception, computer vision, or robotics experience with strong Python and C++ skills.
Machine Learning Engineer owning the full ML lifecycle for multimodal video datasets at Sieve. Fine-tune VLMs, build evaluation/QA pipelines with frontier models, design filtering systems over internet-scale data, and ship production improvements for top AI labs. Requires strong Python, PyTorch, and production ML experience.
Early-career Quantitative Researcher developing ML models and trading signals to predict asset movements and build market-neutral portfolios. Requires STEM degree, Python fluency, and passion for machine learning; embedded in Alpha, Data Science, or Strategy teams.
Build and own evaluation systems, metrics, pipelines, and datasets to rigorously measure the quality of Firecrawl's LLM-ready web data outputs at massive scale, driving model and product improvements.
Lead Data Scientist owning development of core predictive AI/ML models on messy healthcare claims and clinical data to identify waste and improve behavioral health outcomes for payers. Requires 8+ years experience, leadership of data science teams, production model deployment, and strong Python/SQL/ML skills.
Build and operate realtime and batch data pipelines processing billions of events daily at xAI. Design distributed data platforms, own data correctness, create shared datasets for product and business teams, and partner on data acquisition using tools like Spark, Kafka, Flink, and SQL.
Build production ML systems for measuring, predicting, and scaling data quality for frontier AI models. Requires 3-6 years experience in applied ML or related production systems (ranking, recommendations, data quality, fraud) plus strong software engineering skills.
Build analytical and BI infrastructure for Scale's Public Sector unit. Develop scalable data pipelines, models, warehouses, and quality tests from ambiguous processes to enable decision-making; requires 5+ years experience, SQL mastery, Python/R, DBT, and active Secret clearance.
Data scientist partnering with Product, Engineering, and Design to shape mobile consumer creation roadmap at Suno. Design/run experiments, build analyses to identify magic moments driving creative engagement, establish success metrics, and contribute to data foundations. Requires 6-8+ years in consumer tech/mobile, advanced SQL/Python/stats, and end-to-end experimentation experience.
Lead the design and implementation of a unified semantic layer and data models from complex enterprise systems (SAP, Salesforce, Workday) to create AI-ready datasets that power intelligent agents, analytics, and decision-making. Requires 10+ years data engineering experience, semantic modeling expertise, and hands-on AI-generated code deployment.
Develop synthetic sensor simulation models and algorithms using ML techniques like NeRF and Gaussian splatting to generate photorealistic images and realistic lidar/radar data for autonomous vehicles. Requires advanced degree plus 3-5+ years experience, strong ML fundamentals, and Python/deep learning expertise.
Senior Data Scientist on Plaid's Network Value team supporting the Guard product. Translate product questions into analysis, build metrics/OKRs/dashboards, run experiments, and drive data-informed decisions for a 0-to-1 user-facing fintech product.
Field Engineer building and deploying data pipelines, integrations, and backend systems for government customers at customer sites. Requires active Secret clearance, Python, ETL, cloud technologies, and >50% onsite availability.