Skip to content
DatabricksDatabricksSan Francisco, CA

Staff Backend Software Engineer- (AI Platform)

Build and optimize core infrastructure for Databricks Model Serving, focusing on high-throughput, low-latency inference for CPU/GPU workloads. Requires 5+ years in large-scale distributed systems and experience with model serving, inference, routing, autoscaling and observability.

166k – 225k/yr
On-site7+ YOEML Engineering

About the role

Impact

  • Design and implement core systems and APIs that power Databricks Model Serving, ensuring scalability, reliability, and operational excellence.
  • Drive architectural decisions and trade-offs to optimize performance, throughput, autoscaling, and operational efficiency for CPU and GPU serving workloads.
  • Contribute directly to key components across the serving infrastructure — from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling — ensuring smooth and efficient operations at scale.
  • Collaborate cross-functionally with product, platform, and research teams to translate customer needs into reliable and performant systems.
  • Lead technical initiatives that improve latency, availability, and cost-effectiveness across both customer-facing and foundational serving layers.
  • Establish best practices for code quality, testing, and operational readiness, and mentor other engineers through design reviews and technical guidance.

Requirements

  • 5+ years of experience building and operating large-scale distributed systems.
  • Experience in model serving, inference systems, or related infrastructure (e.g., routing, scheduling, autoscaling, and observability).
  • Strong foundation in algorithms, data structures, and system design as applied to large-scale, low-latency serving systems.
  • Proven ability to deliver technically complex, high-impact initiatives that create measurable customer or business value.
  • Experience building architecture for large-scale, performance-sensitive CPU/GPU inference systems.
  • Strong communication skills and ability to collaborate across teams in fast-moving environments.
  • Customer-focused mindset with the ability to align implementation details with product goals.
  • Passion for mentoring, growing engineers, and fostering technical excellence.

Skills

Distributed Systemsmodel servinginference systemsroutingschedulingautoscalingObservabilitySystem DesigncpuGPUAlgorithmsData Structures

Similar roles

ML Engineering jobs
Mozilla

Senior Staff AI & Agentic Systems Engineer

MozillaUnited States

Lead architectural design and implementation of production-grade multi-agent frameworks, LLM orchestration, and autonomous AI agent systems from the ground up at Mozilla's New Products incubator. Requires 7+ years software engineering experience including 2+ years building agentic systems, startup mentality, and deep proficiency with AI tooling.

166k – 260k/yr
Remote7+ YOEML Engineering
Databricks

Staff Software Engineer, AI Search

DatabricksMountain View, CA

As a Staff Engineer for Search, you will build and scale Databricks' next-generation Search product, driving the design and evolution of a highly-performant, cost-efficient, and developer-friendly Search stack. You will also define the long-term vision, mentor senior engineers, and lead strategic efforts.

166k – 225k/yr
On-site10+ YOEML Engineering
Harvey

Mid/Senior/Staff Software Engineer, Agents

HarveySan Francisco, CA +1

Build AI agent systems for legal workflows, optimizing performance via prompt engineering, model selection, tools, and evals. Requires 3+ years experience, Python proficiency, and LLM/agent framework expertise for mid/senior/staff levels.

165k – 312k/yr
On-siteML Engineering
Gusto

Retirement AI Staff Engineer

GustoDenver, CO +2

Staff Engineer building and shipping LLM-powered retirement features end-to-end. Owns architecture decisions, integrates models, and collaborates closely with product and design.

164k – 247k/yr
Hybrid7+ YOEML Engineering
Clear Street

Senior / Staff Software Engineer

Clear StreetUnited States

Build high-performance Rust backend and core AI platform powering an AI-native trading copilot with streaming, low-latency tools, safe trading actions, and robust APIs. Requires 8+ years experience in systems programming, production services, and architecture.

170k – 240k/yr
Remote8+ YOEML Engineering