# Infrastructure Hardware Technical Program Manager

**Company:** [Cerebras Systems](https://hotfix.jobs/companies/cerebras-systems)
**Location:** Unspecified
**Role:** Technical Program Management
**Experience:** 8+ years
**Skills:** Technical Program Management, server architecture, network architecture, cpu/numa, pcie, Linux, firmware, bios, leaf-spine networking, high-performance interconnects, oem/odm management, Risk Management, dependency tracking, AI/ML, hpc
**Posted:** 2026-02-19

> Leads end-to-end delivery of server and network platform programs for AI clusters, coordinating architects, vendors, validation, deployment, and operations. Requires 8+ years of infrastructure-focused technical program management experience and strong knowledge of server architecture, networking, and Linux fleet management.

## Job Description

## Responsibilities
- Own end-to-end program execution for server systems and network equipment in Cerebras clusters, including new platforms, refreshes, and major component/configuration changes.
- Drive requirements gathering and convert inputs into executable plans with clear milestones, readiness gates, and cross-functional deliverables.
- Represent Cluster Architecture in executive reviews, OKR cycles, and leadership/customer forums as needed.
- Build and manage integrated schedules across vendors and internal teams; track dependencies, critical path, and risks.
- Manage OEM/ODM and switch/vendor engagements, including RFI/RFP processes, samples, escalations, and roadmap alignment.
- Partner with Compute, Server Platform, and Network Architects to turn architectural decisions into qualification plans, acceptance criteria, and rollout strategies.
- Lead qualification and release readiness, including lab/staging validation, regression tracking, and go/no-go decisions.
- Own risk and change management into production, including versioning, rollout sequencing, and stakeholder communication.
- Ensure operational readiness with deployment and fleet teams and maintain alignment with rack and physical data center owners on power, cooling, space, and cabling constraints.

## Requirements
- B.S. or M.S. in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience.
- 8+ years in Technical Program Management or similar delivery leadership for server, network, or infrastructure platforms from concept through production.
- Experience coordinating complex server and/or data center network programs across OEMs/ODMs, switch vendors, and internal engineering teams.
- Working knowledge of server architecture, including CPU/NUMA, memory bandwidth, PCIe, NICs, and storage I/O.
- Networking fundamentals, including leaf-spine fabrics, switch platforms, and high-performance interconnects.
- Familiarity with Linux server fleet management, including provisioning, firmware/BIOS, drivers, and field triage.
- Strong multi-team program execution skills, including integrated plans, risk management, dependency tracking, and executive-level communication.
- Ability to operate in ambiguity and keep parallel server and network workstreams aligned.

## Nice to Have
- Experience with AI/ML, HPC, or performance-sensitive distributed infrastructure.

## Compensation and Benefits
- Build a breakthrough AI platform beyond the constraints of the GPU.
- Publish and open source cutting-edge AI research.
- Work on one of the fastest AI supercomputers in the world.
- Enjoy job stability with startup vitality.
- Work in a simple, non-corporate culture that respects individual beliefs.
- Access continuous learning, growth, and support.

## Similar roles

- [Senior Customer Support Operations Readiness Program Manager](https://hotfix.jobs/jobs/b6d6f833-b981-4225-9d87-ddebb7508087) - Mercury - Remote
- [Strategic Projects Lead](https://hotfix.jobs/jobs/76c1beb6-9387-414f-a0b0-9c69a53157a1) - MangoDesk - San Francisco, CA
- [Senior Manager, Quality, Training, and Enablement Operations](https://hotfix.jobs/jobs/73109399-de89-4337-87d8-a920f612e5c6) - Thyme Care - Remote - $140k – $165k/yr
- [Senior Project Manager, Marketing](https://hotfix.jobs/jobs/e18ffd61-9987-4063-bc0b-b5dbf1bc9cd2) - PracticeTek - San Diego, CA - $95k – $100k/yr
- [Technical Deployment Lead, Applied AI](https://hotfix.jobs/jobs/e20b9e3a-0191-4b0c-8607-39f360e1f305) - Anthropic - Austin, TX - $275k – $380k/yr

**Apply:** https://hotfix.jobs/jobs/4225c59f-1c6b-4752-a277-d097c017a35c
**Canonical:** https://hotfix.jobs/jobs/4225c59f-1c6b-4752-a277-d097c017a35c