# Staff Software Engineer - Tools & Infrastructure / DevOps

**Company:** [Cerebras Systems](https://hotfix.jobs/companies/cerebras-systems)
**Location:** Sunnyvale, CA, Toronto, Canada
**Role:** DevOps / SRE
**Experience:** 7+ years
**Skills:** CI/CD, Artifact Repositories, Dependency Management, AWS, Infrastructure Provisioning, Git, Linux, Networking, Kubernetes, Python, Go, Shell, Infrastructure As Code, Observability, Build Systems
**Posted:** 2026-09-01

> Leads the design and operation of CI/CD, build infrastructure, cloud systems, and developer tooling that improve engineering productivity. Requires 7+ years of infrastructure or software engineering experience, strong automation and troubleshooting skills, and technical leadership across teams.

## Job Description

## Responsibilities
- Design, build, and evolve CI/CD pipelines for reliable build, test, and release workflows.
- Own artifact lifecycle systems, including versioning, storage, distribution, dependency management, and reproducible builds.
- Improve code review workflows, branching strategies, repository management, and automated integration processes.
- Provision, monitor, and optimize cloud infrastructure supporting CI workloads for cost, performance, scalability, and reliability.
- Troubleshoot build failures, pipeline bottlenecks, and infrastructure issues; perform root-cause analysis and implement durable fixes.
- Improve build infrastructure, test infrastructure, developer tooling, and automation to increase engineering productivity.
- Identify systemic developer-workflow bottlenecks and lead architectural improvements.
- Contribute to AI tooling that improves engineering productivity and automates repetitive workflows.
- Provide technical leadership, influence engineering standards, and mentor engineers.
- Participate in on-call and incident response for Developer Productivity systems and services.

## Requirements
- 7+ years of professional experience in software engineering, infrastructure engineering, DevOps, developer productivity, or a related area.
- Hands-on experience with CI/CD systems and automated build, test, and deployment infrastructure.
- Experience with artifact repositories, software packaging, dependency management, and reproducible builds.
- Experience with cloud computing platforms and programmatic infrastructure provisioning; AWS preferred.
- Experience with distributed version control, code review workflows, branching strategies, and repository management.
- Strong understanding of Linux/Unix systems, networking fundamentals, and scripting or programming for automation.
- Experience with containerization and container orchestration; Kubernetes preferred.
- Strong troubleshooting and distributed-systems debugging skills.
- Experience leading technical initiatives across multiple teams and improving developer infrastructure at organizational scale.
- Ability to identify architectural bottlenecks, evaluate tradeoffs, and drive long-term improvements.
- Experience operating production systems and participating in on-call, incident response, and postmortem processes.

## Nice-to-haves
- Infrastructure-as-code tools and practices.
- Proficiency in Python, Go, Shell, or another infrastructure-automation language.
- Experience with build systems, build graph optimization, or large-scale build infrastructure.
- Observability experience, including monitoring, logging, alerting, and performance analysis.
- Experience building internal developer platforms, self-service tooling, or developer-facing infrastructure.
- Experience applying AI/LLM tooling to engineering workflows or developer productivity.
- BS/MS in Computer Science or a related field, or equivalent practical experience.

## Similar jobs

- [Staff+ Site Reliability Engineer, Safeguards ML Infra](https://hotfix.jobs/jobs/6492550a-2ff4-498b-8247-470adae7d0c3) - Anthropic - San Francisco, CA - $320k – $485k/yr
- [Staff Infrastructure Engineer](https://hotfix.jobs/jobs/6e086905-4170-407a-a7a4-2e05df0701c5) - Polymarket - New York, NY - $250k – $500k/yr
- [Senior/Staff Infrastructure & Platform Engineer](https://hotfix.jobs/jobs/503b2675-718e-49a4-9692-0ca04a50b707) - Fortanix - Santa Clara, CA - $155k – $230k/yr
- [Staff Network Engineer, App Platform](https://hotfix.jobs/jobs/a410525c-f62d-4037-a6a5-40fa01a905e7) - Scale AI - San Francisco, CA
- [Staff Platform Engineer](https://hotfix.jobs/jobs/d1028610-5698-4ea7-850d-04183a6658da) - Motive - Buffalo, NY - $164k – $236k/yr

**Apply:** https://hotfix.jobs/jobs/2cf4693e-8810-4901-a332-975b49c78512
**Canonical:** https://hotfix.jobs/jobs/2cf4693e-8810-4901-a332-975b49c78512