# Member of Technical Staff, Infrastructure Engineer

**Company:** [Vapi](https://hotfix.jobs/companies/vapi)
**Location:** San Francisco, CA
**Role:** DevOps / SRE
**Salary:** $280k – $314k/yr
**Experience:** 7+ years
**Skills:** SRE, Distributed Systems, Kubernetes, Networking, Cloud Infrastructure, Observability, Incident Response, Failure Analysis, Capacity Planning, Production Automation, Envoy, Postgres, Redis, Kafka, Aurora
**Posted:** 2026-09-12

> Build software, automation, and observability for Vapi’s latency-sensitive distributed infrastructure. The role owns reliability improvements across incident response, capacity, performance, and failure prevention, requiring senior or staff-level experience with SRE and cloud infrastructure.

## Job Description

## Responsibilities
- Learn Vapi’s architecture, production environment, incident history, and reliability practices.
- Build tooling and automation for distributed, latency-sensitive production systems.
- Improve observability, incident response, capacity planning, performance, and production automation.
- Reduce manual operational work and improve detection, diagnosis, and response to failures.
- Own a reliability workstream and become the primary owner for a meaningful part of the reliability surface.
- Deliver durable improvements to failure prevention and recovery.
- Propose roadmaps for future reliability investments.
- Collaborate with Infrastructure and product engineering teams.

## Requirements
- Senior- or staff-level software engineering experience.
- Meaningful SRE, production engineering, or infrastructure experience with distributed systems.
- Ability to write production-quality software and build reliability tooling or automation.
- Deep experience with observability, incident response, failure analysis, capacity, and production reliability practices.
- Experience with Kubernetes, networking, and cloud infrastructure.
- Ability to debug across application and infrastructure boundaries.
- Strong judgment around failure modes and balancing reliability investments with product and engineering velocity.

## Nice-to-haves
- Experience with real-time networking or telephony.
- Experience with Envoy, Postgres, Redis, Kafka, Aurora, ClickHouse, or Google-style SRE practices.

## Compensation and Benefits
- Base salary of $280,000 to $314,000.
- Equity ownership.
- Medical, dental, and vision coverage.
- Flexible time off.
- Quarterly off-sites, catered meals, transportation, gym benefits, and a $10,000 annual learning and development budget.

## Similar jobs

- [Data Center Operations Lead - Partner Site Operations](https://hotfix.jobs/jobs/ef97e148-d0bb-4f9f-b86a-fb28bf724e40) - Anthropic - Austin, TX - $320k – $405k/yr
- [Member of Technical Staff, Release Engineer](https://hotfix.jobs/jobs/fe291883-3543-4e65-807d-413ad6e7d32f) - Vapi - San Francisco, CA - $235k – $264k/yr
- [Software Engineer, Infrastructure](https://hotfix.jobs/jobs/a05a18f3-ba5e-4dc8-af1f-6cdf98a2081b) - Descript - Remote - $220k – $292k/yr
- [Senior Software Engineer - Pipeline Infrastructure & Integration](https://hotfix.jobs/jobs/6480a129-43dc-4fd8-9782-eae819ce9ca2) - Zoox - Foster City, CA - $219k – $315k/yr
- [Senior Infrastructure Engineer](https://hotfix.jobs/jobs/119de60b-8af5-4618-b2f2-b56ff132559b) - Tennr - New York, NY - $200k – $230k/yr

**Apply:** https://hotfix.jobs/jobs/18927189-d9e4-484c-a4f9-e1d2dfc6384d
**Canonical:** https://hotfix.jobs/jobs/18927189-d9e4-484c-a4f9-e1d2dfc6384d