Skip to content
BasetenBasetenSan Francisco, CA

AI Inference Engineer

Build and deploy production AI applications for customers, partnering with their engineering teams from initial problem framing through monitoring and expansion. Requires at least two years of professional experience, production programming expertise with Python preferred, and familiarity with ML model development and deployment.

165k – 330k/yr
Hybrid2+ YOESolutions Architecture

About the role

Responsibilities

  • Develop and maintain production-level software systems and product features using general-purpose programming languages, with a preference for Python.
  • Design, implement, and deploy AI solutions end-to-end, from problem framing and evaluation through production deployment and monitoring.
  • Work directly with customers’ engineering teams across sales, implementation, and expansion.
  • Turn ambiguous objectives into clear specifications and well-defined proofs of concept.
  • Optimize and enhance AI/ML projects and contribute to improvements in the technical stack.
  • Develop features and product requirements with engineering and product teams.
  • Own products and customer projects end-to-end, combining engineering, project management, product management, and technical customer success.
  • Make sound tradeoffs and select appropriate tools while avoiding unnecessary complexity.
  • Demonstrate ownership, accountability, and strong execution.

Requirements

  • Bachelor’s, master’s, or Ph.D. degree in Computer Science, Engineering, Mathematics, or a related field.
  • 2+ years of professional work experience in a fast-paced, high-growth environment.
  • Experience with one or more general-purpose programming languages in a production-level environment; Python is strongly preferred.
  • Familiarity with AI/ML pipelines and the lifecycle of machine-learning model development and deployment.
  • Strong communication skills, particularly when explaining complex technical topics.
  • Experience building or optimizing AI/ML projects is highly valued.

Benefits and Compensation

  • Competitive compensation, including meaningful equity.
  • 100% coverage of medical, dental, and vision insurance for employees and dependents.
  • Flexible PTO, including a company-wide winter break.
  • Paid parental leave.
  • Fertility and family-building stipend through Carrot.
  • Company-facilitated 401(k).
  • Exposure to a variety of ML startups and opportunities for learning and networking.

Skills

PythonMachine Learningartificial intelligenceai/ml pipelinesmodel deploymentproduction softwareProduct ManagementProject ManagementMonitoring
Baseten

Forward Deployed Engineer

BasetenSan Francisco, CA +1

Partners directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform, owning the full journey from exploration to monitoring. Requires 2+ years experience, Python proficiency, and familiarity with AI/ML pipelines in fast-paced environments.

165k – 330k/yrHybrid2+ YOESolutions Architecture
Ramp

Software Engineer, Forward Deployed

RampSan Francisco, CA +1

Designs and builds end-to-end solutions for enterprise customers, collaborating with sales to close deals and drive product roadmap. Requires 2+ years software engineering experience shipping scalable products in fast-paced environment.

168k – 280k/yrHybrid2+ YOESolutions Architecture
Usul

Forward Deployed Engineer

UsulSan Francisco, CA

Forward Deployed Engineer building AI agents, LLMs, and intuitive UIs with React for government technology acquisition workflows and commercial defense contracting platform. Requires 2-8 YOE full-stack experience, confidence presenting to C-suites, willingness to travel to military bases, and passion for national security.

160k – 240k/yrOn-site2+ YOESolutions Architecture
Luma AI

Forward Deployed Engineer

Luma AIRedwood City, CA

Forward Deployed Engineers embed with customers to define problems and build production systems using Luma's AI models and APIs. They own projects end-to-end, from initial conversations to deployed solutions with real data.

170k – 290k/yrHybrid2+ YOESolutions Architecture
Hebbia

Integrations Engineer

HebbiaNew York, NY +1

Build and maintain production-grade data integrations and ingestion pipelines connecting enterprise systems to Hebbia's AI platform. Requires 2-5 years experience with Python, backend services, and API integrations in a high-ownership, on-call production environment.

160k – 265k/yrOn-site2+ YOESolutions Architecture