AI Research Engineer
Builds and deploys production AI features like Notebook Agent for data science workflows, partnering with product teams on experiments, model fine-tuning, and infra. Requires senior AI/ML engineering experience with MLOps, Python/TS proficiency.
About the job
Responsibilities
- Partner with product teams to build AI experiences like the Notebook Agent.
- Run experiments, fine-tune models, deploy AI infrastructure, and build/maintain experimentation tooling.
- Drive Hex's context engine and advance Notebook Agent capabilities (SQL/Python generation, reports, visualizations).
- Build features from 0 to 1.
- Determine architecture and stack for AI-enabled capabilities.
- Ship product experiences changing data science/analyst workflows.
Requirements
- Senior engineer from AI Eng, SWE, or MLE background.
- Experience getting AI/ML capabilities into production for real users.
- Enthusiasm for AI applications to business problems.
- Understanding of core MLOps/SW Architecture for modern ML apps (strong on Infra/MLOps ideal).
- Comfortable in Python & JS/TS.
- Experimentalist mindset with quick iteration and critical thinking.
- Interest in data space, shipping products, empowering users.
- High quality bar for design, correctness, testing.
Our Stack
- Frontend: Typescript, React, Apollo GraphQL, Redux.
- Backend: Typescript, Express, Apollo GraphQL, Postgres, Redis, Kubernetes.
- Infra/CI/CD: Terraform, Helm, AWS.
Compensation
- Salary range: $214,000 - $285,000 (USD), varies by location, skills, experience.
Skills
Python, TypeScript, React, Kubernetes, MLOps, GraphQL, Postgres, Redis, Terraform, AWS
Similar jobs
ML Engineering jobsOptimizes distributed machine learning training and high-throughput offline inference across large accelerator clusters. The role focuses on profiling, scaling efficiency, cluster goodput, GPU performance, and cost-effective processing of autonomy data.
Build and operate production machine-learning systems for search ranking, relevance, extraction quality, and LLM-driven features. The role requires production ML ownership, ranking or relevance expertise, large-scale data experience, Python, and rigorous experimentation skills.
Build and operate production machine-learning systems for content safety, from messy customer data through classification, evaluation, and inference. The role requires 5+ years of ML engineering experience, strong Python and MLOps skills, and sound judgment across classical models and LLMs.
Build AI agent harnesses, models, and product capabilities that enable agents to perform complex work across digital environments. The role combines applied AI research and software engineering, requiring Python proficiency, strong product judgment, and experience with agent tooling, reinforcement learning, or browser technologies.
Builds the platform, verifiers, environments, and grading infrastructure used to evaluate enterprise AI agents at scale. The role combines strong software engineering with expertise in agent runtimes, evaluation design, benchmarks, and production failure analysis.