Skip to content
ZooxZoox

Senior Software Engineer - Pipeline Infrastructure & Integration

Senior engineer owning safety-critical software pipelines and infrastructure, from static and dynamic analysis through CI enforcement, dashboards, and reliability tooling. Requires an advanced technical degree, 7+ years working with large codebases, and expertise in Bazel, Python, backend infrastructure, and C++.

About the job

Responsibilities

  • Own the software safety and stability pipeline ecosystem, including Clang-Tidy, AddressSanitizer, ThreadSanitizer, Harden, and Infer, across weekly, nightly, and continuous integration pipelines.
  • Advance pipeline prototypes into production-grade, CI-enforced systems across branches.
  • Build dashboards and KPIs from issue data to monitor pipeline health and software quality trends.
  • Investigate cross-platform and nondeterministic failures, and develop instrumentation and tooling for root-cause analysis.
  • Support checker rollouts, testing strategy, and infrastructure scaling.
  • Define delivery milestones, automate tooling, report business metrics, participate in code reviews and incident postmortems, and maintain process documentation.

Requirements

  • Master's or PhD degree in computer science, electrical engineering, robotics, aerospace, or a related field.
  • 7+ years of experience working on large codebases.
  • Experience with Bazel and Python.
  • Experience building backend infrastructure, including databases, web servers, and shell scripts.
  • Experience interfacing with C++ and Python software.
  • Ability to create system design diagrams and documentation for customer communication and long-term maintenance.
  • Ability to develop scripts for automation, metrics generation, and traceability.
  • Deep experience with enterprise-scale deployment practices and related issues.

Nice-to-haves

  • Experience with PipeDream, Airflow, and Looker.
  • Proficiency in React, TypeScript, and C++, including C++11, C++14, C++17, memory management, threading, and debugging patterns.

Skills

Bazel, Python, C++, React, TypeScript, Shell Scripting, Databases, Web Servers, CI/CD, Clang-Tidy, Addresssanitizer, Threadsanitizer, Airflow, Looker

Descript

Descript

San Francisco, CA

Software Engineer, Infrastructure
$220k+/yrRemote8+ YOEDevOps / SRE

Own and evolve a broad infrastructure platform spanning cloud, Kubernetes, deployment, reliability, security, and GPU-backed AI systems. The role requires 8+ years operating production distributed systems, strong incident and architecture experience, and practical cloud infrastructure expertise.

The Voleon Group

The Voleon Group

Berkeley, CA
Senior Software Engineer, Developer Experience
$225k+/yrHybrid5+ YOEDevOps / SRE

Build and evolve the developer platform that enables reliable, efficient software delivery across the company. The role requires 5+ years of software engineering experience, strong programming and system-design fundamentals, and expertise in build systems, CI/CD, testing, and deployment automation.

Vapi

Vapi

San Francisco, CA

Member of Technical Staff, Release Engineer
$235k+/yrHybrid7+ YOEDevOps / SRE

Own and improve the CI/CD, testing, and deployment infrastructure that enables fast, safe, observable releases at scale. The role requires strong distributed-systems expertise, hands-on Kubernetes and infrastructure-as-code experience, and a track record of measurable cross-team improvements.

Skydio

Skydio

San Mateo, CA

Senior Software Engineer, Developer Productivity
$200k+/yrOn-site5+ YOEDevOps / SRE

Build and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.

Anyscale

Anyscale

San Francisco, CA

Senior Site Reliability Engineer, Platform Infrastructure
$200k+/yrHybrid5+ YOEDevOps / SRE

Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.