Skip to content
BasetenBaseten

Forward Deployed Engineer

Forward Deployed Engineers own technical outcomes for major AI customers, taking workloads from ambiguous objectives through production while optimizing inference, post-training, and infrastructure. The role requires production software engineering, complex debugging, customer communication, and operational ownership.

About the job

Responsibilities

  • Act as the de facto CTO for customer accounts on Baseten, owning how workloads are designed, operated, and scaled.
  • Translate customer objectives into specifications, proofs of concept, production systems, and success criteria.
  • Design evaluations and benchmarks, diagnose quality or performance gaps, and improve inference, post-training, or evaluation workflows.
  • Respond to mission-critical failures, lead triage and incident resolution, and remain accountable through deployment of fixes.
  • Build tooling, automation, recipes, and reference implementations for evaluation and deployment infrastructure.
  • Influence the product roadmap and ship fixes and features into Baseten's codebase.
  • Manage multiple accounts, prioritize work, and align customers and internal stakeholders on status and risk.

Requirements

  • 1–2 years of software engineering experience shipping and maintaining code in large production systems.
  • Experience debugging complex production issues using logs, metrics, and traces.
  • Ability to own ambiguous technical problems, make decisions under uncertainty, and collaborate with system owners.
  • Interest in working directly with customers and influencing product development.
  • Strong communication skills with both technical and executive audiences.
  • Curiosity about AI inference and training and infrastructure supporting it.
  • Willingness to respond outside regular working hours and participate in an on-call rotation.

Nice-to-haves

  • Depth in infrastructure domains such as storage systems or networking, including InfiniBand or RoCE.
  • Experience operating distributed compute platforms such as Kubernetes, Slurm, or Ray, particularly for GPU workloads.
  • Understanding of LLM architectures and inference engines such as vLLM, TensorRT-LLM, or SGLang.
  • Experience profiling and optimizing GPU workloads for training or serving.
  • Experience with post-training techniques such as SFT and RL.
  • Deep learning experience and fluency with PyTorch or JAX.
  • Operational experience with on-call, incident response, and distributed-systems debugging.

Compensation & Benefits

  • Competitive compensation, including meaningful equity.
  • Medical, dental, and vision insurance fully covered for employees and dependents.
  • Flexible paid time off and company-wide winter break.
  • Paid parental leave.
  • Fertility and family-building stipend through Carrot.
  • Company-facilitated 401(k).
  • Exposure to a variety of machine-learning startups.

Skills

Python, Kubernetes, Slurm, Ray, Gpu Computing, Llm Architectures, vLLM, Tensorrt-Llm, Sglang, PyTorch, JAX, InfiniBand, Roce, Reinforcement Learning, Distributed Systems

Retell AI

Retell AI

San Francisco, CA

Applied AI Engineer
$200k+/yrOn-site2+ YOESolutions Architecture

Build and deploy production Voice AI applications directly with enterprise customers, integrating APIs and SDKs, debugging complex systems, and guiding implementations from proof of concept through launch. The role requires 2+ years of relevant experience, strong programming skills, and excellent customer communication.

Domino

Domino

New York, NY

Forward Deployed Engineer - New Grad
$150k+/yrRemoteSolutions Architecture

New-grad Forward Deployed Engineer developing custom AI solutions and enhancing customer workflows while working directly with clients. Requires programming experience, passion for AI/ML and data science, strong communication skills, and a 2027 degree graduation date.

Retell AI

Retell AI

Redwood City, CA

Forward Deployed Engineer
$150k+/yrOn-siteSolutions Architecture

Build and deploy voice AI solutions and enterprise integrations while working directly with customers and sales. The role is designed for recent graduates with technical internships, coding ability in Python or JavaScript, and an interest in AI and startups.

Standard Template Labs

Standard Template Labs

New York, NY

Forward Deployed Engineer
$140k+/yrOn-site2+ YOESolutions Architecture

Deploy and extend enterprise software in customer environments, building integrations, data pipelines, agent workflows, and tooling. The role requires 2–3 years of software experience, Python and TypeScript familiarity, third-party API integration experience, and comfort working directly with customers.

Greptile

Greptile

San Francisco, CA

Customer Engineer
$140k+/yrOn-site1+ YOESolutions Architecture

Customer-facing technical role partnering with sales to lead discovery, proof-of-concepts, security diligence, and integrations for enterprise AI code-review deployments. Requires at least one year of software or DevOps engineering experience and strong JavaScript/TypeScript fundamentals.