Latest Data & AI jobs
Job results
Conducts cutting-edge research on data for state-of-the-art AI agents like browser and SWE agents, develops prototypes using LLMs and frameworks like PyTorch/JAX, and publishes in top ML venues. Requires 3+ years ML experience and strong cross-functional communication.
Designs, builds, and deploys production-ready AI agents using LLMs, tool use, and reasoning for enterprise problems. Requires 5+ years ML experience, Python proficiency, and Bachelor's in CS/ML/AI.
Develops synthetic data pipelines, production trace agents, and automated agent-building frameworks for enterprise GenAI. Requires 3+ years LLM production experience, top conference publications, and advanced CS degree.
Research Engineer implements and scales post-training techniques like Constitutional AI and RLHF for production AI models, optimizing capabilities, alignment, and safety. Requires strong Python skills, ML systems experience, and ability to handle complex distributed training pipelines.
Research Engineer optimizes and scales production pretraining of frontier AI models, handling performance, debugging, experiments, and on-call incidents. Requires expertise in JAX, TPU, PyTorch, or large-scale ML systems with a 50/50 research-engineering balance.
Software Engineer focused on AI reliability engineering, improving robustness of Claude's serving infrastructure across SDK to accelerators. Partners cross-team on SLOs, monitoring, high-availability systems, incident response, and safeguard models.
Builds next-generation training environments and evaluations for agentic AI models, blending research in reinforcement learning with robust engineering implementation. Requires strong technical judgment, agency, and experience in ML systems.
Software Engineer on Claude Code team builds evaluation systems, tooling, and infrastructure to enhance AI coding capabilities. Collaborates with researchers in fast-paced environment; requires 5+ years experience building complex systems.
Build ML systems to detect and mitigate AI misuse, including classifiers for anomalous behavior, multi-exchange harm monitoring, and agentic safety evaluations. Requires 4+ years ML experience, Python proficiency, and research-to-deployment skills.
Research Engineer focused on mechanistic interpretability, building tools and infrastructure to reverse-engineer neural networks for safer AI. Requires 5+ years software experience, Python proficiency, and AI research contributions.
Designs and optimizes TPU kernels to address performance issues in ML research, training, and inference systems. Provides feedback on model impacts and solves large-scale systems problems, requiring deep accelerator expertise.
Conducts experimental ML research on AI alignment and safety for powerful systems, focusing on scalable oversight, control, and stress-testing. Requires strong software/ML engineering, empirical research experience, and Python proficiency.
Conducts mechanistic interpretability research to reverse-engineer language models, developing methods to understand neural network algorithms for AI safety. Requires scientific research background, Python proficiency, and collaborative engineering mindset.
Research Engineer works end-to-end to remove bottlenecks toward scientific AGI, focusing on long-horizon reasoning, computer use, and model capabilities. Requires 8+ years ML experience, expertise in language model pipelines, distributed systems, and collaborative problem-solving.
Develops next-generation large language models through research, experimentation, and engineering on pre-training team. Requires strong Python/PyTorch skills, ML expertise, and MS/PhD in related field.
Builds large-scale infrastructure for AI scientist training, evaluation, and deployment, resolving bottlenecks in distributed systems for scientific AGI. Requires 6+ years in infrastructure engineering with expertise in ML stacks, containers, and data pipelines.
Research Engineer builds and optimizes reinforcement learning infrastructure for advancing AI capabilities like agentic models, tool use, and reasoning. Requires Python proficiency, ML frameworks experience, and ability to blend research with scalable engineering.
Develops and validates LLM/ML models for healthcare chart abstraction from unstructured/structured data, evaluates cutting-edge NLP/AI techniques for clinical problems, and collaborates with customers to optimize workflows. Requires 5-7+ years data science experience including 3-4 years in healthcare.
Develop and enhance the AI Gateway platform, building unified APIs for AI models with rate limiting, failovers, and integrations. Requires 5+ years experience in JavaScript/TypeScript, backend, and distributed systems.
Develop scalable ML services for data enrichment, managing the full lifecycle from model training with PyTorch/TensorFlow to deploying optimized inference using ONNX/vLLM. Requires 5+ years experience in production ML systems, strong deployment skills, and software engineering proficiency.
Analyzes Attentive's AI products across lifecycle, partnering with PMs and engineers to measure impact via A/B tests, uncover insights, and recommend roadmap changes. Requires 4+ years in analytics, SQL expertise, and stats knowledge.
Develops theories of intelligence grounded in neural network internal structures, focusing on belief geometries in LLMs and biological brains. Conducts experiments bridging mathematics, ML interpretability, and safety research; requires PhD-level quantitative depth and hands-on coding.
Build full stack AI/ML features for finance automation platform, owning greenfield projects with Python/TypeScript/Next.js. Requires 3-5 years full stack exp, 1+ year AI/ML inference, in-person NYC.
Develop backend systems, statistical models, and experiments to optimize Pinterest's ads marketplace, balancing short- and long-term objectives. Requires 10+ years experience, CS/ML degree, and strong engineering/math skills.
Develops ML and optimization models for real-time dynamic pricing and ETA in Lyft's marketplace. Requires MS/PhD in quantitative field, 2+ years algorithms experience, Python proficiency.
Build and improve machine learning models for Pinterest's recommendation systems across Homefeed, Ads, Search, and more. Requires 2+ years experience in ML methods like personalization and recommender systems, plus hands-on work with large-scale data pipelines.
Develop advanced ML models using deep learning to personalize Pinterest experiences across Homefeed, Ads, Search, and more. Requires 4+ years in ML, experience with large-scale systems like Spark/Hadoop, and a degree in CS/ML.
Develop and evolve machine learning technology stack for Pinterest's Ads monetization, building personalized recommendation systems using deep learning, LLMs, and big data tools. Requires 2+ years ML experience in recommender systems, ranking, and large-scale systems.
Develops and evolves machine learning technology stack for Pinterest's Ads monetization, building personalized recommendation systems using deep learning and LLMs. Requires 2+ years ML experience, CS degree, and expertise in large-scale systems.
Senior ML engineer leads development of GenAI models and pipelines for Airbnb's customer support, productionizing at scale. Requires PhD, 10+ years ML experience including 2+ in GenAI, and expertise in NLP, deep learning, and agile AI practices.
Develop and scale generative models like diffusion and flow-matching for autonomous driving plan generation. Collaborate across teams to productize models for real-world deployment, requiring PhD/MSc + 3+ years in generative modeling and strong Python/C++ skills.
New grad Software Engineer for AI Platform teams (Data Platform, Onboard Systems, Technical Infrastructure) at Nuro's self-driving tech company. Build scalable systems for data management, onboard autonomy, and infrastructure using C++, Python, and ML expertise.
Develops state-of-the-art generative models like diffusion and flow-matching for autonomous planning in self-driving tech. Requires PhD or MSc with 2-3 years experience in generative modeling for robotics, strong Python/C++ skills, and top research publications.
Develops scalable ML-based planning and prediction systems for autonomous driving trajectories. Requires expertise in sequential decision making, deep RL, imitation learning, generative modeling, and robotics; M.Sc./Ph.D. preferred with top conference publications.
Research Engineer building evaluations, search systems, and training agentic ML models for tool integration in AI agents. Requires strong research execution, rapid prototyping, and collaboration to productionize ideas.
Lead strategic data science for Reddit's ads platform, owning advanced quantitative solutions in statistics, econometrics, and machine learning to optimize marketplace dynamics, advertiser ROI, and revenue. Requires Master's/PhD with 8-12+ years experience and expertise in ads systems.
Staff Data Scientist guides Reddit's consumer product strategy using data insights, experimentation, and metrics to boost user engagement and retention. Requires advanced degree, 6-10+ years experience, SQL/Python expertise, and causal inference skills.
Data Scientist analyzes customer behavior, product usage, and business performance to drive strategic decisions across Product, GTM, and leadership. Requires 5+ years in data science/product analytics, strong SQL, and experience in high-growth B2B SaaS environments.
Designs and scales distributed data infrastructure for large-scale multimodal AI training and evaluation. Collaborates with researchers to build reliable, high-performance systems in a fast-paced environment.
Builds production AI agents and LLM pipelines for marketing workflows, integrating with data warehouses. Requires strong backend architecture skills, product thinking, and creativity with LLMs; senior role emphasizing impact over years of experience.
Build and scale AI agents automating healthcare workflows. Lead complex feature development using JavaScript/TypeScript/Node.js, LLMs, and DevOps tools. Requires 4+ years experience and CS degree.
Design and deploy low-speed motion planning and control modules for autonomous vehicles, including vehicle-dynamics characterization and optimal-control solutions. The role requires 3+ years of production software experience, hands-on vehicle testing, strong control and numerical-analysis knowledge, and high-performance C++.
Leads development of agentic AI platform to autonomously resolve healthcare insurance claim denials using multi-agent workflows, RAG systems, and LLMs. Requires 5+ years in production ML engineering with Python and frameworks like PyTorch.
Founding Analytics Engineer builds analytics infrastructure from scratch, designs data models and metrics frameworks for healthcare financial data, and creates self-service dashboards. Requires 5+ years experience with SQL, dbt, Python, and cloud data warehouses; NYC-based hybrid role.
Analyzes and reconciles financial and transactional data to improve customer funds reporting, controls, and payments analytics. The role partners cross-functionally with Finance, Compliance, Engineering, and Product while improving dbt pipelines, automation, and data architecture.
Leads design, development, and deployment of AI agents using language models and ML frameworks. Drives scalable, safe AI systems while mentoring teams and collaborating cross-functionally. Requires strong Python skills and AI leadership experience.
Build and improve ML systems and data pipelines infrastructure to support modeling engineers. Requires 2+ years experience, bachelor's in CS/math/sciences, strong coding in Python/Go/Java/C++, and production ML infra expertise.
Builds and maintains end-to-end data pipelines from raw ingestion to analysis, using Python, SQL, and AI tools to create actionable insights for healthcare clients and internal teams. Requires 4+ years experience handling complex data systems with focus on quality and observability.
Staff Data Scientist designs and deploys advanced deep learning models for fraud detection and risk management, leading ML lifecycle and mentoring peers. Requires 8+ years experience, Master's/PhD, and expertise in Python, PyTorch, and diverse data modalities.
Optimizes and builds production inference systems for large language models at scale using PyTorch and high-performance tooling. Requires 3+ years experience in production code, OS concepts, and AI inference systems.