Build and deploy production AI applications for customers, partnering with their engineering teams from initial problem framing through monitoring and expansion. Requires at least two years of professional experience, production programming expertise with Python preferred, and familiarity with ML model development and deployment.
165k – 330k/yrHybrid2+ YOESolutions Architecture
Solution Architect
BasetenSan Francisco, CA
Partners with Sales to translate customer needs into AI technical solutions, leads demos, technical scoping, benchmarking, and POC execution for ML inference deployments across modalities like LLMs and VoiceAI.
165k – 275k/yrOn-siteSolutions Architecture
Forward Deployed Engineer
BasetenSan Francisco, CA +1
Partners directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform, owning the full journey from exploration to monitoring. Requires 2+ years experience, Python proficiency, and familiarity with AI/ML pipelines in fast-paced environments.
165k – 330k/yrHybrid2+ YOESolutions Architecture
Search
Location
3 jobs
Job results
AI Inference Engineer
BasetenSan Francisco, CA +1
Build and deploy production AI applications for customers, partnering with their engineering teams from initial problem framing through monitoring and expansion. Requires at least two years of professional experience, production programming expertise with Python preferred, and familiarity with ML model development and deployment.
165k – 330k/yrHybrid2+ YOESolutions Architecture
Solution Architect
BasetenSan Francisco, CA
Partners with Sales to translate customer needs into AI technical solutions, leads demos, technical scoping, benchmarking, and POC execution for ML inference deployments across modalities like LLMs and VoiceAI.
165k – 275k/yrOn-siteSolutions Architecture
Forward Deployed Engineer
BasetenSan Francisco, CA +1
Partners directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform, owning the full journey from exploration to monitoring. Requires 2+ years experience, Python proficiency, and familiarity with AI/ML pipelines in fast-paced environments.