# Hardware Technical Program Manager, Infrastructure Partner Operations

**Company:** [OpenAI](https://hotfix.jobs/companies/openai)
**Location:** San Francisco, CA
**Role:** Technical Program Management
**Salary:** $226k – $285k/yr
**Experience:** 7+ years
**Skills:** Technical Program Management, infrastructure operations, cloud operations, service delivery, technical account management, slas, KPIs, Dashboards, executive reporting, Incident Management, Root Cause Analysis, AWS, GCP, Azure, hpc
**Posted:** 2026-07-20

> Lead operational delivery and governance for OpenAI's third-party infrastructure partners (cloud providers and compute vendors). Drive SLAs, metrics, escalations, dashboards, and cross-functional programs to ensure reliability for large-scale AI systems. Requires 7+ years in TPM or infrastructure operations.

## Job Description

## Key Responsibilities
- Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, SLAs, and performance expectations.
- Develop operational governance frameworks with strategic partners, including business reviews, operational scorecards, escalation processes, executive reporting, and performance improvement plans.
- Define, track, and continuously improve key operational metrics related to infrastructure availability, deployment execution, incident response, operational health, service quality, and partner performance.
- Build dashboards and reporting mechanisms that provide clear visibility into partner operational performance, risks, trends, and areas requiring executive attention.
- Drive cross-functional coordination between OpenAI teams and external infrastructure providers to resolve operational issues, remove execution blockers, and improve delivery outcomes.
- Lead operational escalations involving infrastructure availability, deployment execution, hardware operations, capacity delivery, or service performance, ensuring timely resolution and clear executive communication.
- Establish repeatable operating rhythms with external partners, including weekly operational reviews, executive business reviews, service reviews, action tracking, and long-term improvement initiatives.
- Partner with Capacity Planning, Hardware Operations, Networking, Deployment, Reliability Engineering, and Supply Chain teams to ensure external infrastructure providers remain aligned with OpenAI’s operational priorities.
- Identify systemic operational risks across partner organizations and proactively drive corrective actions that improve long-term operational effectiveness.

## Qualifications
- 7+ years of experience in Technical Program Management, Infrastructure Operations, Cloud Operations, Service Delivery, or Technical Account Management within large-scale infrastructure environments.
- Experience managing operational relationships with external infrastructure providers, cloud service providers, hardware vendors, or strategic technology partners.
- Strong understanding of hyperscale cloud infrastructure, data center operations, infrastructure delivery, or large-scale distributed systems.
- Experience developing operational KPIs, SLAs, service health metrics, dashboards, and executive reporting for complex technical organizations.
- Demonstrated success leading cross-functional operational programs involving both internal stakeholders and external partners.
- Strong program management skills with the ability to drive accountability across organizations without direct authority.
- Excellent written and verbal communication skills with experience presenting operational performance to senior technical and executive leadership.
- Bachelor's degree in Engineering, Computer Science, Information Systems, Operations, or equivalent practical experience.

## Preferred Skills
- Experience managing cloud infrastructure operations within organizations such as Microsoft Azure, Amazon Web Services (AWS), Google Cloud Platform (GCP), Oracle Cloud Infrastructure (OCI), or other hyperscale cloud providers.
- Experience leading operational governance, service delivery, customer success engineering, technical account management, or infrastructure operations for enterprise cloud customers.
- Strong understanding of service-level agreements (SLAs), operational KPIs, incident management, escalation processes, root cause analysis, and continuous service improvement methodologies.
- Experience building executive dashboards, operational scorecards, business review frameworks, and data-driven performance reporting.
- Familiarity with infrastructure operations supporting GPU infrastructure, AI infrastructure, high-performance computing (HPC), or hyperscale data center environments.
- Experience managing complex cross-company technical relationships while balancing customer priorities, engineering constraints, and operational execution.
- Proven ability to influence senior stakeholders across both internal teams and external partner organizations without direct authority.
- Experience driving continuous operational improvements through metrics, process optimization, and structured governance.

## Similar roles

- [Supply Chain Program Manager (SCPM) - AI Infrastructure](https://hotfix.jobs/jobs/fe7f2ffc-aa73-421a-9efb-02390ab50c65) - OpenAI - San Francisco, CA - $226k – $285k/yr
- [Sr. Technical Program Manager (TPM)](https://hotfix.jobs/jobs/124cf8de-1ade-41f4-ba5f-9aed580aa0d2) - Together AI - San Francisco, CA - $225k – $265k/yr
- [Technical Program Manager, QA Program](https://hotfix.jobs/jobs/ae1892bc-6a33-4e7b-93da-504354053abe) - Fluidstack - Austin, TX - $222k – $295k/yr
- [Infrastructure Delivery Lead](https://hotfix.jobs/jobs/91ee1460-27d8-47fa-8a0f-1efc00990f16) - Fluidstack - New York, NY - $222k – $307k/yr
- [Technical Program Manager Lead, Deployments](https://hotfix.jobs/jobs/6248e3e6-4c18-4924-a2b9-10dce39d8158) - Fluidstack - Austin, TX - $222k – $307k/yr

**Apply:** https://hotfix.jobs/jobs/46db2d73-75a2-4883-9c1f-abefe3d662a8
**Canonical:** https://hotfix.jobs/jobs/46db2d73-75a2-4883-9c1f-abefe3d662a8