Latest AI Research jobs
Job results
Develop AI systems that automate and accelerate research by designing evaluations, building research agents and orchestration infrastructure, and improving model capabilities through training and synthetic data. The role suits strong research or engineering generalists experienced with LLMs, evaluations, agents, infrastructure, or distributed systems.
Develop mathematical models for robots, manipulators, and operating environments while integrating scalable simulation into development and deployment workflows. Requires an advanced degree and experience modeling complex dynamic systems.
Researcher designing and running experiments on chain-of-thought monitorability in frontier LLMs to support scalable oversight and alignment. Requires strong empirical ML expertise with LLMs, deep interest in model behavior/alignment/interpretability, and ability to translate ambiguous questions into concrete experiments.
Conduct machine learning and statistical research to improve unsecured underwriting models, evaluating enhancements through rigorous experimentation and validation. The role requires a graduate degree in a quantitative field, Python modeling experience, and 0–2 years of applied research experience.
Research Advisors apply deep finance, legal, medical, or related expertise to evaluate advanced generative AI systems, shape model-governance frameworks, and collaborate on research and client engagements. Candidates need at least five years of relevant experience, strong analytical skills, and hands-on AI experience.
The fellowship engages experienced software engineers or technical researchers in designing evaluations, datasets, and expert analyses for advanced generative AI systems. Fellows contribute to applied AI research and publications with flexible remote project work.
STEM Fellows apply academic and professional expertise to design evaluation datasets, assess generative AI systems, and contribute research insights and publications. The fully remote, six-month independent contractor opportunity is suited to PhDs, postdoctoral researchers, and professors with relevant domain expertise.
Medical fellows apply clinical expertise to design scenarios, evaluate generative AI decision-making, and provide structured feedback for safer, more accurate healthcare systems. The role requires an MD or DO, board certification, strong clinical reasoning and writing skills, and a relevant medical specialty.
Conducts foundational research on LLMs and multimodal systems, developing architectures, training methods, and optimization techniques and helping transition prototypes into production. The role requires an AI/ML research background, analytical problem-solving, programming experience, and strong research communication.
Conducts empirical economic research on AI’s effects across labor markets, productivity, inequality, and industry transformation. The role develops regional measurement methodologies, leads research collaborations, and translates findings into policy and business insights.
Researches and develops memory and personalization improvements for frontier models through post-training, reinforcement learning, dataset creation, and evaluations. The role requires strong machine-learning expertise, research craftsmanship, and the ability to work across a large codebase with research and product teams.
Conducts research to improve the safety of multimodal AI systems spanning text, vision, and audio. The role requires experience building multimodal models, post-training frontier systems, designing safety evaluations, and translating research findings into reliable model behavior.
Conducts research and develops NLP and LLM-powered capabilities for real-time voice agents, retrieval, and business communications products. The role requires a Master’s or PhD and industry NLP experience, along with Python, PyTorch, and modern LLM expertise.
AI Resident who owns a hard ML/agent problem end-to-end: from proposal and building to evaluation, shipping in production, and rigorous write-up. Requires strong ML fundamentals, Python/PyTorch engineering, and depth in at least one area like post-training, reward modeling, agents, or eval.
Senior AI Research Engineer building and training ML models (LLMs, speech recognition, TTS) to enhance Duolingo's language learning via video calls. Requires advanced ML experience, leadership skills, and an advanced degree or equivalent.
Senior/Staff AI Research Scientist building post-training RL feedback loops and generative models that align frontier biological AI to high-throughput experimental measurements of folding, binding, and function. Requires PhD (or equivalent), hands-on training of models from scratch, and experience with diffusion/transformers/RL.
Principal AI Scientist at Polytope Bio (Astera) to co-direct development of reinforcement learning post-training methods that close the loop between frontier biological AI models and high-throughput experimental data. Requires PhD (or equivalent), hands-on training of large models from scratch, and deep RL expertise; ideal for researchers seeking scientific leadership and potential co-founding role.
PhD research intern role focused on advancing reinforcement learning, machine learning, and foundation models. Work on large-scale training, optimization, inference, long-context tasks, and improving efficiency, reliability, and robustness for real-world AI deployments.
Postdoctoral Young Investigator role developing open multimodal language models (CellOLMo) that integrate single-cell transcriptomics with biological text and knowledge to study Alzheimer's disease and neurodegeneration. Requires PhD in ML, computational biology or related quantitative field plus hands-on transformer model experience.
Sr AI Architect leading Twilio's conversational AI strategy, including memory, knowledge, and behavioral intelligence systems. Requires 15+ years software engineering experience (6+ in production ML at platform scale), deep LLM/LLMOps expertise, and a Master's or PhD in a quantitative field.
Lead a team of research scientists and engineers on GenAI initiatives including evaluation, post-training, agents, and RL. Define multi-year research roadmaps, drive execution from prototype to deployment, publish at top venues, and collaborate cross-functionally in a fast-paced environment. Requires 5+ years research experience, strong publication record, and management background (PhD preferred).
Research Scientist developing cutting-edge 3D vision, reconstruction, and generation models (including foundation models and Gaussian splatting) for autonomous driving and robotics. Requires strong publication record in CV/ML/robotics, MSc/PhD, and experience with PyTorch and computer vision.
Develop machine-learning models and systematic trading strategies by analyzing large financial datasets, researching quantitative finance techniques, and identifying predictive signals. The role requires a STEM degree, Python fluency, and at least two years of systematic trading experience.
Develop novel formal verification methods and hybrid formal engines for AI-generated hardware designs. Collaborate with RTL, ML, and verification teams on model checking, property verification, equivalence analysis, assertion synthesis, and prototyping research ideas on real RTL.
Owns technical delivery of solver pilots and develops production-grade optimisation capabilities for supply chain and operational planning. The role combines mathematical modelling, Python-based solver benchmarking, customer implementations, and collaboration with engineering, data science, and solver vendors.
Conducts research on diffusion, vision-language, and vision-language-action models for autonomous construction robots. The role requires deep-learning R&D experience or advanced graduate training, strong multimodal AI expertise, scalable data workflows, and model optimization for edge deployment.
Build secure infrastructure, control planes, developer tools, and observability for production AI agents. The role requires 6+ years of software engineering experience, strong Python and backend expertise, AWS, distributed systems, AI agent development, and security engineering experience.
Intern building AI prototypes using computer vision and LLMs to automate blueprint reading, cost estimation, and permitting in construction. Requires ownership from research to production, growth mindset, and comfort with unglamorous work.
Build and ship production-grade AI agents, assistants, and reusable skills for security-focused solutions. The role requires hands-on experience with agent frameworks, evaluation, Python, cloud integrations, and containerization, plus a bachelor's or master's degree or equivalent experience.
AI Research Scientist advancing methodological frontiers in healthcare AI at Sprinter Health. Own a research agenda, develop novel architectures/methods, publish at top venues, collaborate with clinicians, and translate findings into production systems. Requires deep ML expertise, strong research taste, and healthcare validation knowledge.
Research Scientist driving fundamental biological discoveries by building large-scale computational analysis pipelines, partnering with experimental biologists on hypothesis-driven research, and leveraging frontier AI models like Claude on petabyte-scale biological data. Requires PhD and end-to-end computational biology research track record with demonstrated breadth.
Develop and integrate custom real-time computer vision and perception algorithms in high-performance C++ pipelines for autonomous systems. The role requires strong computer vision, software engineering, and edge deployment expertise, with experience in areas such as SLAM, tracking, reconstruction, or deep learning preferred.
Develops custom computer vision and perception algorithms in high-performance C++ pipelines for autonomous systems and edge deployment. Requires strong real-time software engineering, computer vision fundamentals, and experience translating algorithms into reliable production systems.
Develops and integrates custom real-time computer vision algorithms and learned perception models into high-performance C++ pipelines for autonomous edge systems. Requires deep C++ and computer vision expertise, with experience in image processing, machine learning, and real-time deployment.
Conduct applied research and engineering to improve language-model behavior in real-time voice conversations. The role focuses on fine-tuning, rigorous evaluation, production failure analysis, data strategies, and safely deploying improvements.
Research-oriented Machine Learning Scientist developing multimodal ML models (NLP, Speech, Computer Vision) for lifelike voice agents. Requires published papers in large-scale deep learning and familiarity with SOTA AI.
Conduct original research on LLM evaluation, routing optimization, and model behavior using billions of real-world generations. Design novel benchmarks, run large-scale empirical studies, and develop statistical foundations for intelligent routing. Requires MS/PhD, publication track record, deep stats/ML expertise, and Python/SQL skills.
Build high-quality, domain-specific benchmarks and infrastructure to rigorously evaluate frontier AI agents on realistic workflows. Requires strong Python/Docker/Linux skills, experience with evals or benchmarks, and a deep understanding of what makes a benchmark reliable and useful.
Lead AI evaluation for Figma's AI-powered products. Define quality metrics, build human + automated eval frameworks (rubrics, golden datasets, LLM-as-judge), manage a small team, and deliver decision-ready insights to Product, Design, and Engineering stakeholders.
Predoctoral Young Investigator role on the AllenNLP team at Ai2, collaborating on open language models and NLP research. Ideal for recent bachelor's/master's graduates preparing for PhD programs, involving mentorship, cutting-edge research, and co-authoring papers.
Research Scientist advancing generative audio models (diffusion/flow matching) for music, focusing on vocal synthesis, post-training alignment (DPO/RLHF), or audio editing. Requires PhD, top publications, and PyTorch expertise to turn research into artist-first Spotify products.
Interns will build open, distributed AI systems and infrastructure across areas such as frontier AI, distributed computing, systems, and cryptography. The role is suited to exceptional builders with substantial project or open-source experience who learn quickly and execute well.
3-month full-time research fellowship at Base Labs focused on open-source LLMs and frontier AI. Fellows receive 1:1 senior mentorship, $15k stipend, full support, and a path to full-time roles while producing publishable research from the San Francisco office.
Conduct hands-on post-training research on LLMs including RL, distillation, and routing models. Collaborate with customers, labs, and engineering to turn techniques into production products and shape the research agenda. Requires proven research background in post-training LLMs and ability to ship impactful work.
Design, implement, and evaluate novel ML models for computational pathology to predict patient outcomes. Requires PhD in ML, computer vision or statistics, strong publication record, and expertise in PyTorch.
Design, implement, and evaluate novel self-supervised foundation models for clinical multi-modal data and precision medicine. Requires PhD in ML/statistics, strong publication record in top venues, and expertise in PyTorch/deep learning.
Member of Technical Staff generating clinical insights from multi-modal AI models for precision medicine in oncology. Requires MD/PhD, lead authorship on high-impact papers, ML knowledge, and Python skills.
Research Scientist developing novel diffusion models and generative algorithms for Large Tabular Models (LTMs) on enterprise data. Requires PhD and strong track record in generative ML; experience with diffusion models and PyTorch/JAX preferred.
Conduct frontier security research across code, cloud, and AI technologies, developing threat models, agentic workflows, and functional prototypes that enable contextual understanding and new product capabilities. Requires 5+ years of security or security research experience and strong systems architecture expertise.
Lead applied research applying formal methods, automated reasoning, and AI (incl. LLMs) to improve correctness, reliability and velocity of Snowflake's cloud data platform and distributed systems. Translate ideas into production capabilities; requires PhD + 8+ years experience and track record of impact.