Latest Data & AI jobs
Job results
Develops and deploys cutting-edge AI models and LLM-based solutions for internal tools and customer products. Collaborates with product/engineering teams to prototype, iterate, and shape technical direction in a fast-paced startup environment. Requires 3+ years experience with Python, TypeScript, and AI frameworks.
Designs, builds, and maintains ML training and serving infrastructure, providing support to research teams. Requires 4+ years in ML infrastructure, cloud platforms like Kubernetes and Google Cloud, and GPU experience.
Leads GPU inference engineering for Sora, optimizing model serving efficiency, kernel-level performance, and scalability. Collaborates with research and product teams to build reliable infrastructure for multimodal AI models.
Develops advanced LLM-based knowledge graphs, RAG techniques, agents, and NLP features to enhance Onyx's AI knowledge retrieval platform. Requires 3+ years AI/ML experience with PyTorch/TensorFlow and strong software engineering skills.
Develops quantitative models and trading strategies for futures markets using financial data, machine learning, and statistical analysis. Requires 2+ years experience, STEM degree, and Python proficiency.
Builds and operates Habitat, OpenAI's core online database platform handling high-QPS, latency-sensitive workloads. Owns end-to-end distributed systems for storage, caching, routing, CDC, and privacy; requires 8+ years experience with Rust/Python expertise.
Builds ambitious RL environments and evaluation systems to measure and steer frontier AI models toward safe AGI. Requires strong ML research engineering, statistical skills, and red-teaming mindset for end-to-end project ownership in fast-paced setting.
Designs, develops, and deploys scalable ML systems and backend infrastructure for healthcare applications, translating LLM research into production while handling large-scale clinical data. Requires 3+ years ML backend experience, 5+ years software development, and backend languages like Python.
Develops novel AI agent applications using language models, manages model alpha program with OpenAI, architects risk AI workflows, and conducts experiments to evaluate model capabilities for B2B SaaS products.
Builds advanced AI systems using massive-scale distributed machine learning to achieve breakthrough performance. Requires strong programming skills and experience with large distributed systems.
Build and deploy retrieval, ranking, classification, and LLM-based systems that improve large-scale search quality. The role requires deep search or recommender-systems expertise and at least five years of relevant project experience.
We are seeking full-time researchers to develop and analyze learning algorithms for a system that represents domain knowledge as modular probabilistic models. The role involves theoretical problems with immediate implementation applicability in finance and scientific research.
Designs, develops, and deploys AI-driven agentic systems and integrates emerging AI research to enhance platform capabilities for software organizations. Requires 2+ years AI/ML experience, proficiency with LLMs, and Bachelor's/Master's in CS or related field.
Develop machine-learning models and systematic trading strategies from large financial datasets, conducting quantitative-finance research and simulations. The role requires a STEM degree, Python proficiency, strong problem-solving skills, and interest or experience in alpha research.
Builds and maintains scalable data processing pipelines and backend systems for a data curation platform that optimizes training data for ML models. Partners with researchers to integrate research capabilities, ensuring reliability and security for customer data.
Designs evaluation frameworks and benchmarks to test AI agents' autonomy, reasoning, and reliability in data pipelines and warehouses. Requires experience in LLM benchmarking, reinforcement learning, Python, PyTorch/JAX, and data engineering tools.
Build and own backend infrastructure for an AI-powered wealth manager, including multi-agent LLM systems, financial-data platforms, and agentic financial-planning capabilities. Requires 5+ years of backend engineering experience and proficiency with Node.js, TypeScript, Python, PostgreSQL, AWS, and distributed-systems technologies.
Develops RL environments and fine-tunes language models using PPO, DPO, and KTO to enhance agentic capabilities for data infrastructure tasks. Requires deep RL expertise, LLM fine-tuning knowledge, and strong problem-solving skills.
Build and maintain anti-abuse and content moderation infrastructure to ensure AI safety. Collaborate with engineers on AI alignment techniques, incident response, and risk mitigation using Python and cloud tools.
Optimizes and extends ML model serving infrastructure for LLMs, speech, and vision models, focusing on high-throughput, low-latency inference using frameworks like VLLM and SGLang. Requires deep PyTorch expertise, systems programming, and performance engineering for reliable production deployment.
Staff Engineer architects scalable agentic AI systems for property management platform, integrating LLMs to automate workflows. Requires hands-on expertise in agentic frameworks, React/Python/Node.js, and leading AI feature development in startup environment.
Builds and optimizes distributed training infrastructure for large-scale multimodal AI models across thousands of GPUs. Requires deep expertise in PyTorch, CUDA, parallelization techniques, and GPU clusters.
Develops and deploys ML models for NLP, retrieval, ranking, reasoning, dialog, and code-generation systems. Requires Master's/PhD, 2+ years experience with production ML, deep NLP expertise, Python, and frameworks like PyTorch/TensorFlow.
As an Autonomy Engineer, you will develop, integrate, and test core path planning capabilities for aerial platforms. This role involves writing software for real autonomous aircraft systems and collaborating with DoD experts to integrate autonomy software onto OEM hardware.
Build and optimize scalable infrastructure for training large frontier AI models, bridging research and production. The role requires strong software engineering, Python and ML framework expertise, and hands-on experience with distributed training at scale.
Optimizes training performance for advanced language models by developing scalable software, GPU kernels, distributed training systems, and profiling tools. The role requires strong software engineering skills, Python and ML framework proficiency, and experience with CUDA or Triton.
Develops LLM-powered tools to boost developer productivity 10x, including code assistance, automated reviews, and task automation for autonomy software. Requires 7+ years experience as full-stack/backend developer fluent in Python or C++ with LLM optimization knowledge.
Optimizes large AI models for high-volume, low-latency production and research environments. Collaborates with researchers and engineers on inference stack performance, requiring 5+ years experience with PyTorch, GPUs, CUDA, and distributed systems.
Pioneers post-training techniques to enhance LLMs for agentic systems, focusing on tool-use, continuous updates, synthetic data infrastructure, and capability evaluations. Requires Python/PyTorch proficiency, post-training expertise, and proven research impact.
Data Scientist embedding with product teams to define metrics, run A/B tests, build dashboards, and drive data-informed decisions for consumer and enterprise AI products. Requires 5+ years quantitative experience with SQL/Python in hyper-growth environments.
Designs and implements core Python framework components for building real-time AI agents that see, hear, and speak. Owns features end-to-end with strong API design skills and experience in production Python systems.
Builds and productionizes AI models and systems for life sciences document generation, focusing on LLMs, NLP robustness, and reliable deployment pipelines. Bridges ML research, software engineering, and product needs in a high-stakes domain.
Designs, develops, and deploys ML models focused on fine-tuning multimodal LLMs for fraud detection and application automation in financial profiles. Requires 3+ years experience with Python, PyTorch, and expertise in information extraction from financial documents.
Build and ship LLM-powered product features, evaluation frameworks, and retrieval systems end-to-end. The role requires production experience with multi-provider LLM solutions, large-context architectures, and TypeScript, React.js, and Node.js, with regular in-person work in London.
Develops advanced AI agents for healthcare revenue recovery, focusing on human-like conversational AI, model improvements, LLM orchestration, and evaluation frameworks to handle insurance interactions and billing tasks.
Owns end-to-end ML lifecycle from prototyping clinical prediction models to productionizing and deploying them using production-grade Python/SQL. Requires PhD +3yrs or Master's +5yrs experience with MLOps tools like SageMaker/MLFlow.
Engineers optimize ML systems for performance at scale, focusing on GPU utilization, inference engines, and container runtime to boost throughput and reduce latency for language and diffusion models. Requires 5+ years experience with PyTorch, CUDA, and performance debugging.
Research Engineer working on frontier audio AI models, including training, post-training, data curation, architectural improvements, and rigorous evaluation. Requires at least 3 years of AI experience and evidence of solving challenging machine learning problems through projects or research.
Builds state-of-the-art document processing infrastructure using LLMs, including QA agents, optimizers, multimodal models, and self-correcting systems. Monitors production models, runs experiments, and owns large product areas for real-world customer impact.
Develops computational operators to enable LLMs to perform transparent, verifiable reasoning over thousands of iterations on knowledge states like scientific papers. Requires LLM experience, reasoning intuitions, and strong software engineering skills.
Designs and develops cutting-edge multimodal AI systems integrating text, speech, and vision. Conducts research on representation learning using Python, JAX, PyTorch, TensorFlow, with expertise in distributed training and autoregressive models.
As a Software Engineer on the Behavior Capabilities team, you will develop and implement algorithmic advancements to expand the robot's driving abilities in complex scenarios, focusing on improving trip progress and vehicle uptime.
Hands-on AI Engineer prototyping and refining LLM-based features, upgrading capabilities with prompt engineering and fine-tuning, while creating evaluations and migration processes for Fathom's meeting AI product. Requires Python proficiency, analytics skills, and Master's degree.
Leads development of ML algorithms for robot perception, including scene understanding, tracking, segmentation, and multi-modal foundation models using sensor data. Requires deep expertise in deep learning, computer vision, and production ML pipelines, collaborating across autonomy teams.
Part-time contractor responsible for ingesting and quality-checking retailer data, communicating issues, and providing limited ticket triage and customer support. Requires reliable weekend and early-morning availability, attention to detail, and comfort with spreadsheets.
Develops deep learning models using imitation and reinforcement learning to generate safe, efficient driving trajectories for autonomous vehicles. Collaborates with Perception, Planning, and Simulation teams; requires ML expertise, Python fluency, and transformer experience.
Build and scale production AI systems while researching and experimenting with novel modeling ideas. The role requires strong software engineering, Python and ML framework proficiency, GPU kernel development, distributed training experience, and familiarity with Transformer-based sequence models.
Builds scalable data pipelines and infrastructure for AI research, processing petabyte-scale anime data across 10k GPUs. Partners with researchers using distributed systems, big data tools, cloud services, requires 3+ years generalist experience.
Research Engineer owning datasets for training world simulation AI models, designing multimodal datasets, running experiments, and building data pipelines to enhance model capabilities across tasks like robotics and creative tools. Requires 4+ years in ML with experience in generative models and frameworks like PyTorch or JAX.
Drive strategic data projects end-to-end, building models and metrics that shape go-to-market strategy and operations. Requires 4+ years analytics experience or advanced quantitative degree, strong SQL/Python skills, and ability to work directly with leadership on high-impact business decisions.