Skip to content

Director, Data Center Operations

Lead design, fit-out, and commissioning of data center sites focused on power, cooling, and IT infrastructure for high-density GPU workloads. Build and manage a 20-person operations team, oversee multi-site portfolio, and establish processes from scratch.

About the job

Responsibilities

  • Own the design, fit-out, and commissioning of white space sites across the US and Asia, with a focus on power distribution (PDUs), cooling distribution (CDUs), and IT-adjacent infrastructure
  • Build and lead a ~20-person break-fix and smart hands team from scratch — define the operating model, hire the initial team, and establish the processes and playbooks that keep sites running
  • Manage a portfolio of 5+ sites across two regions in various stages of deployment and live operation
  • Partner with vendors, contractors, and equipment suppliers to drive site deployments to schedule and quality
  • Establish operational standards, runbooks, and escalation processes for a nascent but rapidly growing infrastructure function
  • Serve as the technical authority on data center infrastructure decisions — from white space evaluation through to live production operation
  • Travel to Asia periodically to oversee regional site deployments and support the local team

Requirements

  • Deep, hands-on technical knowledge of data center power and cooling systems at the IT-adjacent layer — PDUs, CDUs, power distribution, and white space fit-out
  • Proven experience designing and commissioning data center infrastructure, from evaluation through to live operation
  • Experience operating data center infrastructure at meaningful scale, supporting large production workloads
  • People leadership experience — you've hired, developed, and led technical operations teams
  • A builder's instinct — you're comfortable standing up teams and functions from scratch, writing the playbook where none exists, and operating with ambiguity
  • Strong vendor and contractor management skills — you know how to hold external partners accountable to timeline and quality

Nice to Have

  • Experience with GPU or AI infrastructure deployments, multi-site or multi-region portfolio management, or familiarity with Asian markets

Compensation

US base salary range: $250,000 - $300,000 + equity + benefits

Skills

Pdus, Cdus, Power Distribution, Cooling Systems, White Space Fit-Out, Data Center Infrastructure, Gpu Infrastructure, AI Infrastructure, Vendor Management, Runbooks

Fluidstack

Fluidstack

United States

Principal Operations Engineer, Network
$258k+/yrRemote7+ YOEDevOps / SRE

This principal-level role owns operational excellence for a hyperscale AI data center network fleet, leading readiness, high-risk changes, audits, and incident resolution across sites. It requires extensive mission-critical network operations experience, routing and optical networking expertise, and 50–75% travel.

Snowflake

Snowflake

Menlo Park, CA

Principal Software Engineer - Performance Engineering
$264k+/yrOn-site12+ YOEDevOps / SRE

Leads Snowflake’s cloud infrastructure performance strategy by evaluating new hardware, building benchmark and validation systems, and translating performance data into pricing, capacity, and rollout decisions. Requires 12+ years in performance, systems, or infrastructure engineering and deep cloud hardware expertise.

Cloudflare

Cloudflare

Atlanta, GA
Principal Systems Engineer, DevTools
$200k+/yrHybrid7+ YOEDevOps / SRE

Build and operate AI-powered developer tools, internal MCP integrations, and platform capabilities across the engineering organization. The role requires strong coding and debugging skills, Kubernetes operations experience, and the ability to lead projects, improve developer experience, and mentor teammates.

Fluidstack

Fluidstack

Remote

Principal Operations Engineer, Mechanical
$150k+/yrRemote10+ YOEDevOps / SRE

As a Principal Operations Engineer, Mechanical, you will be the senior technical authority for mechanical and cooling infrastructure across hyperscale AI data centers. You will lead site assessments, drive operational readiness, review designs, and ensure precision execution of critical systems.

AllSpice

AllSpice

Boston, MA
Principal / Staff / Senior Infrastructure Engineer
No salary listedHybrid8+ YOEDevOps / SRE

Own and scale secure cloud infrastructure, deployments, observability, compliance, and incident response for a hardware collaboration platform. The role requires substantial cloud or security engineering experience, AWS and Linux expertise, and the ability to lead cross-functional infrastructure initiatives.