Staff Platform Engineer building scalable infrastructure to support low-latency voice AI product used hundreds of times daily by millions. Architect core platform components connecting product and ML, anticipate scaling bottlenecks, improve developer experience, and set technical direction for growing team.
270k – 350k/yr
On-site7+ YOEDevOps / SRE
About the role
What You'll Do
Build the seam between product and ML: credentials, proxy, routing, experimentation.
See the next scaling bottleneck two quarters early and fix it before anyone notices it existed.
Architect and own whole parts of the platform end to end.
Make every product engineer faster — dev experience, deploy speed, abstractions.
Help set technical direction as the team goes from three to eight.
Requirements
You've rewritten an entire service because you realized the bugs were all coming from bad architecture.
You've seen systems break, understood why, and now think twelve months ahead.
You build the foundation patiently so that everything after it ships faster.
You get excited about complex problems with lots of constraints and edge cases.
You care as much about the user journey as the technical requirements.
Tech Stack
Python/FastAPI
PostgreSQL
Redis
ClickHouse
Temporal/SQS
Terraform
AWS
gRPC
ALB
Fargate
Benefits
Equity on employee-friendly terms, including early exercise and extended exercise windows with a combination of ISOs and RSUs.
Health, dental and vision. We cover 100% of the premium on the Anthem Blue Cross Gold PPO, and 75% of premiums for dependents. (US)
401(k) with up to a 4% match. (US)
16 weeks of paid parental leave (12 weeks for non-birthing parent).
Flexible time off and company holidays.
All meals in the office, plus commuter benefits and relocation reimbursement if you’re moving to San Francisco.
A high-spec laptop, monitor, and a whisper-friendly microphone.
Staff Software Engineer building foundational multi-cloud platform infrastructure for Astronomer's Astro DataOps platform. Requires deep distributed systems expertise, Kubernetes operator-level knowledge, strong Go proficiency, and experience driving technical strategy at scale.
275k – 377k/yrHybrid7+ YOEDevOps / SRE
Senior Staff Software Engineer, Infrastructure
TemporalUnited States
Designs and implements large-scale public cloud infrastructure, builds complex distributed systems and microservices. Requires 10+ years experience, expert skills in performance tuning, concurrency, multiple cloud providers like AWS/GCP/Azure, and graduate degree or equivalent.
260k – 325k/yrRemote10+ YOEDevOps / SRE
Member of Technical Staff, AI Reliability & Monitoring Engineering Lead
PostmanSan Francisco, CA
Lead AI reliability engineering for Postman's API and agentic systems, building monitoring, observability, and automation for high availability. Requires strong SRE/DevOps background in large-scale AI infrastructure and cloud platforms.
256k – 276k/yrHybridDevOps / SRE
Member of Technical Staff, AI Platform & Architecture (Infrastructure)
PostmanSan Francisco, CA +3
Builds and maintains distributed AI infrastructure for model training, inference, and data pipelines. Requires experience in GenAI systems, distributed computing, Python/Go, and scaling AI workloads on GPUs/cloud.
256k – 276k/yrHybridDevOps / SRE
Staff Site Reliability Engineer
EarninMountain View, CA
Lead EarnIn's AI-first reliability engineering strategy. Define SLOs/SLIs, build AI agents for incident response and on-call automation, and partner with engineering teams to embed AI-assisted operations across production systems on AWS.