Latest Data & AI jobs
Job results
Build and operate a declarative data platform supporting ingestion, transformation, and consumption. The role requires cloud data platform, orchestration, infrastructure-as-code, event-driven architecture, SQL, and distributed systems expertise.
Build and scale analytical data platforms, warehouses, and pipelines supporting customer dashboards and high-volume systems. The role requires 3+ years of backend and infrastructure experience plus expertise in data architecture, schema design, integrations, and scalable data platforms.
Builds foundational ML platform infrastructure including model serving pipelines, GPU scheduling systems, and CI/CD for large-scale multimodal AI models. Requires 5+ years in distributed systems with expertise in Python, Kubernetes, and AWS.
Build and scale high-performance software and distributed reinforcement-learning systems for LLM post-training. The role combines production engineering, research tooling, algorithm optimization, and collaboration across infrastructure and scientific teams.
Develops and optimizes diffusion models and generative AI for fashion design software. Conducts research, implements techniques for design quality and controllability, and deploys to production. Requires Master's/PhD, 3+ years experience, PyTorch proficiency.
Build, scale, and maintain complex AI/NLP models for the Verneek AI platform, focusing on production deployment of large-scale systems. Requires 3+ years Python/PyTorch experience and BSc in CS; NLP expertise preferred.
Develops methods for AI agents to self-improve post-training through prompt optimization, continual learning from long-horizon tasks, hypothesis testing, and scalable experiments. Requires expertise in LLMs, agent frameworks, and impactful research track record.
Staff Machine Learning Engineer responsible for architecting and operating production-grade ML systems, NLP and Generative AI applications, and end-to-end MLOps on AWS. The role requires 8+ years of experience, strong software engineering fundamentals, and expertise in observability and scalable model deployment.
Develops, deploys, and optimizes production AI systems using LLMs, NLP, and machine learning tools for quant research and market applications. Requires at least two years of hands-on AI or deep learning experience and familiarity with production model deployment.
Builds and improves AI agents and workflows for automating enterprise operations like IT, HR, and finance. Requires 4+ years software engineering with applied AI focus, full-stack skills in TypeScript/Node.js/React, and strong product sense.
The Data Operations Analyst ensures operational data accuracy through quality assurance, SQL-based analysis, process documentation, and cross-functional issue resolution. The role requires a bachelor's degree or equivalent experience, strong attention to detail, and comfort working an overnight schedule.
Advances AI coding models through research, experimentation, and system optimization on the Codex team. Collaborates to improve code generation, reasoning, and performance for real-world deployment.
Build scalable voice interface features through ML prototyping, infrastructure development, and inference optimization. Requires Python fluency, LLM expertise, and startup experience.
Develops scalable voice interface features through prototyping, infrastructure building, and R&D. Requires PhD in ML or related field, top conference publications, Python/LLM fluency, and strong engineering skills.
Build and maintain core data models, pipelines, and reporting infrastructure using modern data stack tools to enable decision-making across product, ops, and GTM teams. Requires 3-6 years of production data/analytics engineering experience and expert SQL.
Develops and trains large-scale diffusion models for image/video generation, controllability modules like IPAdapters/ControlNets, and novel research techniques for production. Requires proven experience with image/video models at scale and deep learning frameworks.
Develops low-latency, high-throughput inference services for OCR and multimodal models, optimizing batching, kernels, and autoscaling while evaluating serving frameworks. Requires 3+ years in performance engineering or ML systems with strong Python and GPU experience.
Develops and fine-tunes specialized vision and language models for document understanding, including OCR, layout, tables, and charts. Requires 3+ years ML experience with PyTorch/JAX and strong engineering focus; onsite in San Francisco.
Develops and optimizes transformer-based vision-language models for physical security, owning full-cycle training, fine-tuning, and deployment optimization. Collaborates cross-functionally to integrate models into the platform using PyTorch/TensorFlow and advanced AI techniques.
Develops controllability, personalization, and productization for video foundation models using fine-tuning, RL, and evaluation techniques. Requires deep expertise in visual generative models, PyTorch, and product-focused research for creative workflows.
Develops state-of-the-art agentic AI systems for issue triage, debugging, and predictive analytics using Sentry's error datasets. Requires 4+ years experience (with MS/PhD) or 6+ (with Bachelor's), Python, PyTorch expertise, and production ML deployment skills.
Curates structured datasets on companies, people, and networks using investigative research in English and Chinese sources. Partners with engineering for data quality and supports customers on supply chain intelligence, requiring 4+ years in investigations or due diligence.
Builds AI systems for voice-first interfaces, audio intelligence pipelines, and real-world conversation analysis using Python/TypeScript and ML tools like PyTorch/OpenAI. Deploys production LLM systems with customer focus in NYC office.
Data Scientist shapes infrastructure scaling for OpenAI's AI models and products by building datasets, metrics, forecasting/optimization models, and dashboards. Partners with engineering, research, and product teams; requires 5+ years experience, SQL/Python expertise.
Builds and debugs AI agent infrastructure for healthcare automation, including prompt engineering, runtime issue tracing, evaluation datasets, simulation tooling, data pipelines, and observability dashboards. Requires 2-7 years experience with production LLMs/AI agents and TypeScript proficiency.
Builds state-of-the-art vision models and novel LLM techniques for document processing infrastructure. Monitors production models, runs experiments, and owns large product areas in a high-impact founding role.
Builds production AI agents and workflows for governance, risk, and compliance, fine-tuning LLMs on proprietary data for high-accuracy tasks like regulatory analysis and risk assessment. Requires 3+ years applied AI experience with production ML systems emphasizing explainability and responsibility.
Leads research on post-training data curation for foundation models, designing algorithms to generate/improve instruction and preference datasets, and unifying pre/post-training optimization. Requires 3+ years deep learning research, post-training experience with vision/language/multimodal models, and PyTorch proficiency.
Conducts machine learning research in medical NLP for conversation summarization, evidence extraction, and outcome prediction. Publishes at top AI conferences, deploys models to production, and requires MS/PhD plus strong PyTorch/TensorFlow experience.
Builds AI-powered product features and backend services for intelligent workflows in Linear's product development platform. Collaborates with product/design to integrate foundation models, optimize prompts, and architect reliable AI infrastructure using React, GraphQL, Node, and Temporal.
Bridge research and production to build agentic LLM and NLP systems for finance and legal workflows. Own experiments combining latest research with customer use cases while embedding in the full software development lifecycle.
Analyzes user behavior, product usage, and AI model performance to guide product decisions and improve automation pipelines. The role requires a bachelor's degree, 1–2 years of AI analyst experience, Python and SQL proficiency, and statistical analysis skills.
Leads the architecture and development of production AI products and a centralized AI platform for accounting automation. The role requires deep Python and backend engineering expertise, scalable system design, and experience with LLM applications, RAG, integrations, and technical leadership.
Build and own core AI agent features and SDKs for an autonomous identity platform. Work across the stack (React, GraphQL, Python/Flask) and mentor engineers as the team scales.
ML Research Engineer trains foundational embedding models for web search, designs novel transformer architectures, creates datasets and evals, and beats state-of-the-art performance. Requires graduate-level ML expertise and ability to implement transformers in PyTorch.
Builds and improves large-scale pretraining data pipelines, mixtures, and curation methods for Cohere’s language models. The role combines software engineering and research, requiring Python, data-pipeline development, and experience with large datasets and processing frameworks.
Staff AI Software Engineer builds and scales AI systems including PromptQL AI assistant, data engine, and runtime infrastructure for enterprise reliability. Requires 6+ years experience with AI/ML, distributed systems, LLMs, Python fluency, and technical leadership.
Builds and owns data pipelines, ETL processes, and infrastructure to power reporting and decision-making. Requires 4+ years experience with Python, PostgreSQL, Snowflake, and data modeling.
Junior Research Scientist develops and refines ML architectures and LLMs for conversational AI in housing/healthcare, applies advanced math/ML to predict behaviors and optimize interactions. Requires PhD in math/physics/CS or related, strong ML foundation, and onsite work.
Build and improve production generative AI and multi-agent features across the full stack, using LLMs to solve product problems. The role requires strong programming experience, hands-on LLM prompting or GenAI implementation, and a product-oriented approach.
Develops embedding models and retrieval systems to enable frontier AI models to access relevant information dynamically. Requires deep expertise in representation learning, vector retrieval, and transformer LLMs, with experience scaling ML systems.
Conducts research in mechanistic interpretability to understand deep network representations and enhance AI safety. Requires PhD or equivalent in ML/CS, 2+ years research engineering with Python, and passion for safe AGI.
Build and post-train frontier AI models at scale, bridging research and production through scalable training software, distributed infrastructure, and performance optimization. The role requires strong software engineering skills and experience with large-model training and post-training.
Designs, develops, and maintains ML systems for content recommendations, search, spam detection, and labeling using Bluesky's social graph. Requires 3+ years in ML/data science focused on recommendations/search, Python/PyTorch proficiency, and rapid experimentation skills.
Leads architecture and evolution of a large-scale data platform spanning ingestion, lakehouse storage, streaming, governance, and access. Requires 10+ years of software engineering experience, deep production expertise with Spark and distributed systems, and strong cloud data-platform experience.
Build and optimize high-performance inference infrastructure for OpenAI's multimodal models handling image, audio, and other non-text inputs at scale. Collaborate with research and product teams on low-latency production systems using GPU workloads and inference tooling.
Develops cutting-edge perception models using fused sensor data for autonomous vehicles, focusing on joint detection/tracking and embeddings for downstream systems. Requires MS/PhD in CS, PyTorch experience, and production deployment skills.
Advances AI capabilities through cutting-edge reinforcement learning research, training intelligent agents for models like o1 and o3. Requires RL research background, strong coding skills, and ability to iterate quickly in a fast-paced environment.
Designs and tests zero-knowledge proof systems for biometric likeness verification in privacy-focused identity solutions. Requires expertise in zk-SNARKs/STARKs, cryptography, and prototyping tools like Circom and Semaphore.
Build and deploy scalable NLP and LLM solutions for voice agents, classification, information retrieval, and agentic applications. The role requires 3+ years of machine learning and NLP experience, strong Python/PyTorch skills, and experience with production model deployment.