# Data Center Compute Infrastructure

**Company:** [OpenAI](https://hotfix.jobs/companies/openai)
**Location:** San Francisco, CA
**Role:** DevOps / SRE
**Salary:** $230k – $490k/yr
**Experience:** 5+ years
**Skills:** Distributed Systems, gpu clusters, high performance computing, AI Infrastructure, data center development, mechanical engineering, electrical engineering, power systems, Networking, facilities engineering, manufacturing, supply chain, cloud scale platforms
**Posted:** 2026-07-16

> Build, scale, and operate OpenAI's global compute infrastructure for frontier AI models like GPT-5.6. Solve complex cross-disciplinary problems spanning distributed systems, hardware, ML infrastructure, power/cooling, manufacturing, supply chain, and data center development at unprecedented scale.

## Job Description

## Key Responsibilities
- Help build, scale, and operate OpenAI’s global compute infrastructure.
- Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
- Improve the reliability, performance, efficiency, and scalability of critical infrastructure.
- Partner with cross-functional teams to bring new compute capacity online quickly and reliably.
- Identify bottlenecks across technical, operational, and physical systems, and develop practical solutions.
- Build tools, processes, systems, or infrastructure that improve execution at scale.
- Contribute to the long-term architecture and operational maturity of OpenAI’s compute footprint.

## Qualifications
- Experience building, scaling, or operating complex technical systems.
- Enjoy working on ambiguous, high-impact problems where the path forward is not always defined.
- Comfortable collaborating across disciplines, including software, hardware, operations, and physical infrastructure.
- Strong technical judgment and a bias toward execution.
- Care deeply about reliability, speed, safety, and operational excellence.
- Excited by the challenge of building infrastructure at unprecedented scale.
- Work directly supports the development and deployment of frontier AI.

## Preferred Skills
- Experience with AI infrastructure, high-performance computing, distributed systems, GPU clusters, or cloud-scale platforms.
- Worked on hardware systems, manufacturing, supply chain, data center development, or large capital infrastructure projects.
- Domain expertise in civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering.
- Helped bring new technical platforms, data centers, factories, or large-scale systems from concept to production.
- Experience operating in fast-moving environments where technical depth and execution speed both matter.

## Similar roles

- [Software Engineer, Productivity - Inference Runtime](https://hotfix.jobs/jobs/69dc68e4-afb2-408a-99b3-d0a75f5b59a5) - OpenAI - San Francisco, CA - $230k – $385k/yr
- [Software Engineer, Core Network Engineering](https://hotfix.jobs/jobs/edc46bbe-94a3-427d-8e51-12997465af0c) - OpenAI - San Francisco, CA - $230k – $342k/yr
- [Software Engineer, Productivity - Model Performance](https://hotfix.jobs/jobs/02edb47f-3a80-4c92-97af-12456fa045ea) - OpenAI - San Francisco, CA - $230k – $385k/yr
- [Software Engineer, Productivity - Networking](https://hotfix.jobs/jobs/6c5db7b0-2fe1-4065-86b8-431a296ad080) - OpenAI - San Francisco, CA - $230k – $385k/yr
- [Software Engineer, Compute Infrastructure](https://hotfix.jobs/jobs/35add999-67ee-4269-8975-f19937c751aa) - OpenAI - San Francisco, CA - $230k – $405k/yr

**Apply:** https://hotfix.jobs/jobs/062a5299-b497-4412-9f02-9f7ba1672d72
**Canonical:** https://hotfix.jobs/jobs/062a5299-b497-4412-9f02-9f7ba1672d72