Skip to content
BasetenBaseten

AI Engineer

Build and ship agentic AI product experiences, internal automation, and customer-facing features across the stack. The role requires 5+ years of software engineering experience, hands-on experience with AI or LLM-powered products, Python proficiency, and strong autonomy.

About the job

Responsibilities

  • Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to general availability.
  • Design harnesses, execution flows, and guardrails that make AI systems reliable in production.
  • Build internal automation and AI tooling that measurably increases the velocity of engineering, research, and go-to-market teams.
  • Partner with research engineers to turn internal research workflows into customer-facing features.
  • Define and instrument evaluations to measure output-quality improvements.
  • Work across the API layer, backend, agent orchestration, and frontend to implement end-to-end features.
  • Use the company's training and inference products to develop intuition around customer workflows.
  • Identify and replace manual processes with AI-driven solutions.
  • Fix bugs and resolve customer issues with urgency.

Requirements

  • 5+ years of experience building and shipping software applications.
  • Demonstrated experience building AI- or LLM-powered products, agents, or agentic workflows used by real users.
  • Strong software engineering fundamentals.
  • Ability to understand how systems, models, and harnesses work under the hood.
  • Proficiency in Python and fluency in at least one other programming language.
  • Ability to work autonomously in a fast-moving environment with limited structure.
  • Ability to move between customer-facing product work and internal tooling and automation.
  • Strong communication skills and ability to bridge technical depth with business needs.

Nice-to-haves

  • Experience as a founding engineer or early startup employee.
  • Experience building evaluations, agent observability, or tooling for non-deterministic systems.
  • Familiarity with LangChain, Claude Code, Codex, OpenCode, or MCP.
  • Experience with supervised fine-tuning, reinforcement learning, synthetic data generation, LoRA, or full fine-tunes.
  • Frontend fluency.

Compensation and benefits

  • Competitive compensation, including meaningful equity.
  • 100% coverage of medical, dental, and vision insurance for employees and dependents.
  • Flexible PTO, including a company-wide winter break.
  • Paid parental leave.
  • Fertility and family-building stipend through Carrot.
  • Company-facilitated 401(k).
  • Exposure to a variety of ML startups and related learning and networking opportunities.

Skills

Python, LangChain, Claude Code, Codex, Opencode, Mcp, Supervised Fine-Tuning, Reinforcement Learning, Synthetic Data Generation, Lora, Frontend Development, Agent Orchestration, Evaluation, API Development, Backend Development

Mercor

Mercor

San Francisco, CA

Research Scientist, APEX Benchmarks
$200k+/yrOn-siteAI Research

Leads the design, measurement, publication, and adoption of APEX benchmarks evaluating frontier models on economically valuable professional work. The role requires rigorous research judgment, strong coding and statistical skills, and excellent communication across technical, commercial, and research audiences.

Tessera Labs

Tessera Labs

San Jose, CA

Research Scientist
$200k+/yrOn-siteAI Research

Research Scientist defining and executing research on reliable long-horizon agents in enterprise environments. The role focuses on post-training and reinforcement learning, agent memory, evaluation, verification, and structured representations, combining hands-on experimentation with product delivery and publication.

OpenAI

OpenAI

San Francisco, CA

People Research Scientist
$198k+/yrOn-siteAI Research

Conduct rigorous people research and applied data science to evaluate talent programs, organizational health, and employee experiences. The role requires advanced expertise in research design, experimentation, measurement, causal inference, statistical modeling, and responsible handling of sensitive employee data.

The Voleon Group

The Voleon Group

New York, NY
Member of Research Staff, Causal Inference
$250k+/yrHybridAI Research

Conducts causal inference research for financial market prediction and portfolio optimization, developing and validating models from research through live trading. Requires Ph.D.-level coursework, strong causal inference and statistics expertise, mathematical ability, and production Python skills.

Earnin

Earnin

Mountain View, CA

Software Engineer (Gen AI)
$181k+/yrHybrid3+ YOEAI Research

Build agent-driven chatbots and generative AI workflows for financial-wellness products, owning features from design through impact measurement. The role requires at least three years of software engineering experience, strong system design, maintainable coding practices, and a bachelor’s degree or equivalent experience.