Skip to content
OpenAIOpenAI

Data Center Compute Infrastructure

Build, scale, and operate OpenAI's global compute infrastructure for frontier AI models like GPT-5.6. Solve complex cross-disciplinary problems spanning distributed systems, hardware, ML infrastructure, power/cooling, manufacturing, supply chain, and data center development at unprecedented scale.

About the job

Key Responsibilities

  • Help build, scale, and operate OpenAI’s global compute infrastructure.
  • Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
  • Improve the reliability, performance, efficiency, and scalability of critical infrastructure.
  • Partner with cross-functional teams to bring new compute capacity online quickly and reliably.
  • Identify bottlenecks across technical, operational, and physical systems, and develop practical solutions.
  • Build tools, processes, systems, or infrastructure that improve execution at scale.
  • Contribute to the long-term architecture and operational maturity of OpenAI’s compute footprint.

Qualifications

  • Experience building, scaling, or operating complex technical systems.
  • Enjoy working on ambiguous, high-impact problems where the path forward is not always defined.
  • Comfortable collaborating across disciplines, including software, hardware, operations, and physical infrastructure.
  • Strong technical judgment and a bias toward execution.
  • Care deeply about reliability, speed, safety, and operational excellence.
  • Excited by the challenge of building infrastructure at unprecedented scale.
  • Work directly supports the development and deployment of frontier AI.

Preferred Skills

  • Experience with AI infrastructure, high-performance computing, distributed systems, GPU clusters, or cloud-scale platforms.
  • Worked on hardware systems, manufacturing, supply chain, data center development, or large capital infrastructure projects.
  • Domain expertise in civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering.
  • Helped bring new technical platforms, data centers, factories, or large-scale systems from concept to production.
  • Experience operating in fast-moving environments where technical depth and execution speed both matter.

Skills

Distributed Systems, Gpu Clusters, High Performance Computing, AI Infrastructure, Data Center Development, Mechanical Engineering, Electrical Engineering, Power Systems, Networking, Facilities Engineering, Manufacturing, Supply Chain, Cloud Scale Platforms

Perplexity

Perplexity

San Francisco, CA
Member of Technical Staff
$220k+/yrRemote4+ YOEDevOps / SRE

Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.

Firecrawl

Firecrawl

San Francisco, CA

Cloud DevOps Engineer
$240k+/yrHybrid5+ YOEDevOps / SRE

Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.

Fireworks AI

Fireworks AI

San Mateo, CA
Member of Technical Staff - Reliability Engineering
$240k+/yrHybrid5+ YOEDevOps / SRE

Owns reliability standards, incident management, observability, failure testing, and automation for a high-throughput AI infrastructure platform. The role requires deep Linux, networking, software, cloud-native, and distributed-systems experience, along with the ability to influence teams across the organization.

Tessera Labs

Tessera Labs

San Francisco, CA

AI Platform Engineer
$200k+/yrRemote5+ YOEDevOps / SRE

Build and own production-grade AI agent infrastructure across multiple clouds, with responsibility for Kubernetes, Terraform, observability, security, reliability, and automation. Requires 5+ years of cloud infrastructure experience and strong CI/CD, networking, and production operations expertise.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.