Latest Data & AI jobs
Job results
Senior/Staff AI Research Scientist building post-training RL feedback loops and generative models that align frontier biological AI to high-throughput experimental measurements of folding, binding, and function. Requires PhD (or equivalent), hands-on training of models from scratch, and experience with diffusion/transformers/RL.
Principal AI Scientist at Polytope Bio (Astera) to co-direct development of reinforcement learning post-training methods that close the loop between frontier biological AI models and high-throughput experimental data. Requires PhD (or equivalent), hands-on training of large models from scratch, and deep RL expertise; ideal for researchers seeking scientific leadership and potential co-founding role.
Build and own data foundations powering model training, product development, and analytics at Parallel. Design scalable ingestion pipelines, storage/serving layers, quality/lineage/observability systems, and anticipate scaling needs in a fast-growing AI infrastructure company.
The Senior Database Administrator will administer AWS-hosted PostgreSQL databases, lead query and schema optimization, guide data architecture, and mentor engineers. The role requires 7+ years managing enterprise databases, strong SQL expertise, and experience with cloud database services and AI-assisted technical workflows.
Lead the architecture and development of Dialpad's autonomous Agentic AI platform, building multi-agent orchestration, memory systems, real-time reasoning, and tool execution for enterprise workflows. Requires 10+ years experience, technical leadership at Staff/Principal level, and deep expertise in LLM platforms, agent frameworks, and production AI infrastructure.
Lead the architecture and development of Dialpad's autonomous Agentic AI platform, building multi-agent orchestration, memory systems, real-time reasoning, and tool execution for enterprise workflows. Requires 10+ years experience, prior technical leadership at Staff/Principal level, and deep expertise in LLM platforms, agent frameworks, and production AI infrastructure.
Staff Marketing Technology Engineer who owns a BigQuery/dbt data platform and builds secure AI applications and agents on Google Cloud. Requires 5+ years of data engineering experience plus strong Python, SQL, cloud deployment, software engineering, and data governance expertise.
Founding Data Analyst on the FinOps team building the company's data foundation, operating scorecard, and dashboards. Requires 3-5 years experience, strong SQL/Python, comfort with messy data, and executive presence in a high-impact, cross-functional role at an AI-native healthcare startup.
Lead the data quality team at HUDHUD to build QC systems, validation methods, and experiments that measure and improve training data for frontier AI agents and RL environments. Requires deep data quality intuition, Python/Docker/Linux proficiency, and experience turning research insights into production evaluation pipelines.
Develops and optimizes petabyte-scale cloud database systems, focusing on high-performance data processing, query optimization, and scalability solutions. Requires 2+ years experience, fluency in Java or C++, strong CS fundamentals, and onsite work in Menlo Park or Bellevue.
PhD research intern role focused on advancing reinforcement learning, machine learning, and foundation models. Work on large-scale training, optimization, inference, long-context tasks, and improving efficiency, reliability, and robustness for real-world AI deployments.
Design and build core infrastructure for AI inference on Cloudflare's global network of GPUs and accelerators. Optimize scheduling, routing, reliability, and observability for low-latency, serverless LLM and model serving at the edge. Requires expert Rust and distributed systems experience.
Lead a 6-10 person data science team for Duolingo's User Growth pillar. Own experimentation, forecasting, BI, and analytical strategy to drive acquisition, retention, and business decisions. Requires 5+ years managing data scientists, graduate degree, and expertise in ML, causal inference, and SQL/Python.
Lead development of state-of-the-art vision, VLM, and VLA models for autonomous perception systems. Own challenging ML problems, build data pipelines and deployment infrastructure for production autonomy, and mentor engineers while bridging research to real-world systems. Requires 7+ years experience, strong CV/ML expertise, and C++/Python proficiency.
Develop and deploy advanced vision, VLM, and VLA machine learning models for autonomous systems perception. Own models from training through optimization and deployment on embedded hardware, collaborating with research and engineering teams to deliver production capabilities for complex real-world environments.
Develop and deploy state-of-the-art vision, VLM, and VLA models for autonomous perception systems. Requires 7+ years experience (with PhD), mastery of ML and computer vision, PyTorch/TensorFlow, TensorRT/ONNX, C++/Python, and ability to obtain SECRET clearance.
Project Controls Analyst responsible for cost management, tracking, forecasting, performance monitoring, variance analysis, and reporting on construction and engineering projects. Requires 2-5 years in cost controls or financial analysis and advanced Excel skills.
Own commercial data science end-to-end for Headway's GTM and payer business. Lead strategic analyses, forecasting, and causal inference for the Chief Commercial Officer and executive stakeholders while technically leading other data scientists.
Staff Data Scientist partnering with product, risk, finance, and operations teams to drive strategy through predictive insights, advanced experimentation, causal inference, and ML-based fraud interventions in a marketplace setting. Requires 7+ years data science experience, strong SQL/Python skills, and expertise in statistical modeling and risk models.
Supports data entry, data quality, database maintenance, and analysis while assisting customers with product configuration, demonstrations, onboarding, and feedback. The role suits a detail-oriented, curious candidate with Excel and data-framework experience, ideally in fixed income or leveraged credit.
Lead the unified Data & AI engineering function at BuildOps. Own data platforms, pipelines, ML infrastructure, governance, and a high-performing team to power AI products and establish BuildOps as the trusted system of record for commercial contractors. Requires 10+ years experience scaling data teams and deep expertise in modern data/ML platforms.
Senior Data Platform Engineer owning data infrastructure for identity and fraud detection products. Build scalable ETL/ELT pipelines, data observability, and storage layers using Python/Golang, Spark, AWS, and databases. 5+ years experience required; mentor juniors and collaborate with product/DS teams.
Build and scale Addepar's Analytics Platform using the Data Lakehouse to optimize data analytics, workflows, Ops, and Data Governance. Requires 5+ years experience in platform development for data engineering outcomes, strong Java/Python skills, cloud platforms, CI/CD pipelines, IaC, and observability tools.
Postdoctoral Young Investigator role developing open multimodal language models (CellOLMo) that integrate single-cell transcriptomics with biological text and knowledge to study Alzheimer's disease and neurodegeneration. Requires PhD in ML, computational biology or related quantitative field plus hands-on transformer model experience.
Senior Manager leading a new strategic health analytics function to deliver rigorous, evidence-based analysis supporting payer and employer conversations. Hands-on role combining deep US healthcare expertise (payer/provider), advanced analytics, customer-facing narrative development, team building, and mentoring.
Lead and grow a data science team to drive insights, experimentation, and predictive modeling that improve provider experience metrics and product engagement at Headway. Requires 8+ years data science experience including 3+ years managing teams, strong product and cross-functional partnership skills, and technical fluency in SQL, Python/R, and causal inference.
Build and optimize Cerebras’s production GPU inference stack across APIs, vLLM, PyTorch, ROCm, distributed systems, and AMD infrastructure. The role requires 8+ years of software engineering experience, strong C++ and Python skills, and deep expertise in GPU performance, reliability, and model serving.
Senior Data Analyst on the Data Engineering team who bridges client needs, internal healthcare product knowledge, and data-driven outcomes. Requires deep healthcare claims expertise, expert SQL and AI/LLM usage for analysis and automation, payment integrity experience, and collaboration with clients and internal teams.
GTM Data Scientist building predictive churn, look-alike, and marketing mix models to measure campaign impact, identify expansion opportunities, and drive data-informed decisions for Sales, Marketing, and RevOps at a rapidly growing beverage tech company.
Build foundational AI agent infrastructure at Rippling, owning agent creation, invocation, skill packaging, plugin connectivity, and automations. Lead platform and distributed systems work with 8+ years experience, technical leadership, and cross-functional impact.
Build scalable AI platform infrastructure including agent orchestration frameworks, evaluation systems, and high-performance serving layers to accelerate AI development and deployment at Mixpanel and for customers. Requires 2+ years software engineering experience with hands-on LLM/agent integration and full-stack fundamentals.
Serve as the analytical engine for consumer business decisions at a privacy-first mobile carrier. Combine market research, business intelligence, financial modeling, and customer insights to drive pricing, planning, and strategy; present recommendations to executives. Requires 4+ years in consulting, BI, or market research with strong SQL/Excel and analytical skills.
Build and operate scalable backend systems, data pipelines, storage solutions, and architectures supporting healthcare advertising products, analytics, and machine learning workflows. The role requires at least one year of data engineering experience, cloud proficiency, and strong Java, SQL, and Python skills.
Senior MLOps Engineer building scalable ML infrastructure and pipelines for autonomous battery-electric rail vehicles. Lead design of distributed training, experiment tracking, deployment, and monitoring systems for safety-critical perception and autonomy models. Requires 5+ years building large-scale systems with 2+ years in ML infrastructure.
Senior Software Engineer building automated evaluation pipelines, test infrastructure, and monitoring systems to validate quality of Deepgram's speech, audio, LLM, and multimodal AI models before production release. Requires 5+ years building test/evaluation frameworks, strong analytical skills, and backend experience in Python/Rust/Go.
Develop metrics and large-scale evaluation pipelines to assess fidelity of Zoox's GenAI-powered autonomous vehicle simulator. Requires 5+ years in quantitative evaluation or robotics, strong Python/data analysis skills, and statistics knowledge.
Build and maintain data pipelines for reconciling card-network settlements against internal transaction data to power accurate financial, tax, and regulatory reporting at scale for a global payments platform. Requires 3+ years experience with batch/event-driven pipelines, orchestration tools, and strong SQL/programming skills in a high-stakes financial environment.
Build and scale inference infrastructure for generative audio models including TTS, voice conversion, and ASR. Design high-performance, low-latency serving systems using Kubernetes, CI/CD, and GPU optimization to bridge research and production.
Lead the Sales Analytics team at Gusto to rebuild data foundations, evolve forecasting with causal inference and advanced methods, drive revenue insights, and shift the team toward strategic statistical thinking in an AI-first environment. Requires 10+ years data science experience including 4+ years leading teams, deep Salesforce and SaaS revenue domain expertise, and hands-on technical skills.
Build and operate reliable, scalable infrastructure and automation for OpenAI's research workloads and data systems (acquisition, processing, ingest, search). Requires strong systems and distributed systems experience, Kubernetes, Linux, networking, and software engineering to improve reliability and reduce operational toil.
Build and deploy production-grade AI agents for high-stakes fraud, risk, compliance, and AML decisions at enterprise customers. Own the full agent lifecycle from scoping and tool engineering through evals, tuning, and platform impact. Requires strong LLM experience, software engineering fundamentals, and customer-facing skills.
Build actuarial models (total-cost-of-care, PMPM, MLR) from claims data to quantify Sprinter Health's long-term economic impact on payers. Partner with health plan actuaries, support commercial pricing, and ensure rigorous causal measurement of in-home care interventions.
Build and ship AI agent orchestration workflows, integrations, and shared infrastructure that power automated growth across paid media, lifecycle, organic, CRO, incentives and more at Kraken. Requires 3+ years building agentic systems or automation with LLM APIs, strong API integration skills, and growth/marketing context.
Build and maintain production ML systems including training/inference pipelines, model serving via APIs/batch, monitoring for drift, and automated retraining. Productionize models from prototypes with strong Python, MLOps, and reliability focus.
Lead Garner Health Intelligence as Director of Applied Science, owning doctor-ranking algorithms, data assets, and the enterprise Insights platform. Build and manage teams of scientists, researchers, and PMs while driving both technical innovation and commercial business results in healthcare.
Develop and productionalize ML and optimization models for Lyft's pricing and ETA systems. Requires advanced degree, production ML/algorithms experience, and Python proficiency to solve large-scale marketplace problems.
Staff Applied Data Scientist owning end-to-end pricing experimentation, modeling, and strategy for Coinbase's consumer and business products. Requires 8+ years in data science focused on causal inference and pricing, plus strong SQL/Python/R skills.
Build and improve production AI agent systems for OpenAI's GTM workflows. Own the end-to-end improvement loop using feedback, evaluation, experimentation and backend services to drive measurable gains in customer engagement, pipeline and team productivity. Requires 4+ years building reliable LLM-powered production systems plus strong product judgment.
Staff Software Engineer driving reinforcement learning infrastructure for Claude's coding capabilities at Anthropic. Design APIs/frameworks, embed with research teams to build and hand off maintainable systems, improve research code reliability, and ensure production RL run health. Requires deep Python expertise, API design track record, and failure-mode intuition.
Build and own Python frameworks, APIs, and infrastructure for Anthropic's RL environments and agent runtimes. Embed with research teams to productionize their work, design for correctness in stateful distributed systems, and create self-service tooling for production debugging.