Skip to content

Staff Software Engineer - Tools & Infrastructure / DevOps

Leads the design and operation of CI/CD, build infrastructure, cloud systems, and developer tooling that improve engineering productivity. Requires 7+ years of infrastructure or software engineering experience, strong automation and troubleshooting skills, and technical leadership across teams.

About the job

Responsibilities

  • Design, build, and evolve CI/CD pipelines for reliable build, test, and release workflows.
  • Own artifact lifecycle systems, including versioning, storage, distribution, dependency management, and reproducible builds.
  • Improve code review workflows, branching strategies, repository management, and automated integration processes.
  • Provision, monitor, and optimize cloud infrastructure supporting CI workloads for cost, performance, scalability, and reliability.
  • Troubleshoot build failures, pipeline bottlenecks, and infrastructure issues; perform root-cause analysis and implement durable fixes.
  • Improve build infrastructure, test infrastructure, developer tooling, and automation to increase engineering productivity.
  • Identify systemic developer-workflow bottlenecks and lead architectural improvements.
  • Contribute to AI tooling that improves engineering productivity and automates repetitive workflows.
  • Provide technical leadership, influence engineering standards, and mentor engineers.
  • Participate in on-call and incident response for Developer Productivity systems and services.

Requirements

  • 7+ years of professional experience in software engineering, infrastructure engineering, DevOps, developer productivity, or a related area.
  • Hands-on experience with CI/CD systems and automated build, test, and deployment infrastructure.
  • Experience with artifact repositories, software packaging, dependency management, and reproducible builds.
  • Experience with cloud computing platforms and programmatic infrastructure provisioning; AWS preferred.
  • Experience with distributed version control, code review workflows, branching strategies, and repository management.
  • Strong understanding of Linux/Unix systems, networking fundamentals, and scripting or programming for automation.
  • Experience with containerization and container orchestration; Kubernetes preferred.
  • Strong troubleshooting and distributed-systems debugging skills.
  • Experience leading technical initiatives across multiple teams and improving developer infrastructure at organizational scale.
  • Ability to identify architectural bottlenecks, evaluate tradeoffs, and drive long-term improvements.
  • Experience operating production systems and participating in on-call, incident response, and postmortem processes.

Nice-to-haves

  • Infrastructure-as-code tools and practices.
  • Proficiency in Python, Go, Shell, or another infrastructure-automation language.
  • Experience with build systems, build graph optimization, or large-scale build infrastructure.
  • Observability experience, including monitoring, logging, alerting, and performance analysis.
  • Experience building internal developer platforms, self-service tooling, or developer-facing infrastructure.
  • Experience applying AI/LLM tooling to engineering workflows or developer productivity.
  • BS/MS in Computer Science or a related field, or equivalent practical experience.

Skills

CI/CD, Artifact Repositories, Dependency Management, AWS, Infrastructure Provisioning, Git, Linux, Networking, Kubernetes, Python, Go, Shell, Infrastructure As Code, Observability, Build Systems

Anthropic

Anthropic

San Francisco, CA
Staff+ Site Reliability Engineer, Safeguards ML Infra
$320k+/yrHybrid8+ YOEDevOps / SRE

Staff-level site reliability engineer responsible for safely deploying and operating safeguards infrastructure across model releases and cloud platforms. The role emphasizes production change management, high-stakes incident response, and automating manual launch and validation processes.

Polymarket

Polymarket

New York, NY

Staff Infrastructure Engineer
$250k+/yrOn-site7+ YOEDevOps / SRE

Staff Infrastructure Engineer responsible for designing and operating scalable infrastructure for growth systems, including onboarding, referrals, and user acquisition. The role requires 7+ years of production infrastructure experience, strong reliability instincts, and independent judgment in a high-autonomy environment.

Fortanix

Fortanix

Santa Clara, CA

Senior/Staff Infrastructure & Platform Engineer
$155k+/yrOn-site7+ YOEDevOps / SRE

Leads the architecture, development, and operation of cloud, Kubernetes, on-premises, and hybrid infrastructure, while building developer platforms and CI/CD automation. Requires at least six years of infrastructure or related engineering experience, deep Kubernetes expertise, strong programming skills, and technical leadership.

Scale AI

Scale AI

San Francisco, CA

Staff Network Engineer, App Platform
No salary listedOn-site7+ YOEDevOps / SRE

Own the network architecture and standards for a multi-cloud enterprise AI platform deployed across Kubernetes environments and customer-controlled networks. The role requires deep cloud and Kubernetes networking expertise, strong security fundamentals, and the judgment to establish scalable, supportable connectivity patterns.

Motive

Motive

Buffalo, NY
Staff Platform Engineer
$164k+/yrOn-site7+ YOEDevOps / SRE

Staff Platform Engineer will build and improve automated delivery pipelines, developer environments, infrastructure, and release systems across the engineering organization. The role requires 6+ years of engineering experience, a bachelor’s degree, and expertise with CI/CD, cloud infrastructure, containers, and infrastructure as code.