Director, Data Center Operations
Lead design, fit-out, and commissioning of data center sites focused on power, cooling, and IT infrastructure for high-density GPU workloads. Build and manage a 20-person operations team, oversee multi-site portfolio, and establish processes from scratch.
About the job
Responsibilities
- Own the design, fit-out, and commissioning of white space sites across the US and Asia, with a focus on power distribution (PDUs), cooling distribution (CDUs), and IT-adjacent infrastructure
- Build and lead a ~20-person break-fix and smart hands team from scratch — define the operating model, hire the initial team, and establish the processes and playbooks that keep sites running
- Manage a portfolio of 5+ sites across two regions in various stages of deployment and live operation
- Partner with vendors, contractors, and equipment suppliers to drive site deployments to schedule and quality
- Establish operational standards, runbooks, and escalation processes for a nascent but rapidly growing infrastructure function
- Serve as the technical authority on data center infrastructure decisions — from white space evaluation through to live production operation
- Travel to Asia periodically to oversee regional site deployments and support the local team
Requirements
- Deep, hands-on technical knowledge of data center power and cooling systems at the IT-adjacent layer — PDUs, CDUs, power distribution, and white space fit-out
- Proven experience designing and commissioning data center infrastructure, from evaluation through to live operation
- Experience operating data center infrastructure at meaningful scale, supporting large production workloads
- People leadership experience — you've hired, developed, and led technical operations teams
- A builder's instinct — you're comfortable standing up teams and functions from scratch, writing the playbook where none exists, and operating with ambiguity
- Strong vendor and contractor management skills — you know how to hold external partners accountable to timeline and quality
Nice to Have
- Experience with GPU or AI infrastructure deployments, multi-site or multi-region portfolio management, or familiarity with Asian markets
Compensation
US base salary range: $250,000 - $300,000 + equity + benefits
Skills
Pdus, Cdus, Power Distribution, Cooling Systems, White Space Fit-Out, Data Center Infrastructure, Gpu Infrastructure, AI Infrastructure, Vendor Management, Runbooks
Similar jobs
DevOps / SRE jobsThis principal-level role owns operational excellence for a hyperscale AI data center network fleet, leading readiness, high-risk changes, audits, and incident resolution across sites. It requires extensive mission-critical network operations experience, routing and optical networking expertise, and 50–75% travel.
Leads Snowflake’s cloud infrastructure performance strategy by evaluating new hardware, building benchmark and validation systems, and translating performance data into pricing, capacity, and rollout decisions. Requires 12+ years in performance, systems, or infrastructure engineering and deep cloud hardware expertise.
Build and operate AI-powered developer tools, internal MCP integrations, and platform capabilities across the engineering organization. The role requires strong coding and debugging skills, Kubernetes operations experience, and the ability to lead projects, improve developer experience, and mentor teammates.
As a Principal Operations Engineer, Mechanical, you will be the senior technical authority for mechanical and cooling infrastructure across hyperscale AI data centers. You will lead site assessments, drive operational readiness, review designs, and ensure precision execution of critical systems.
Own and scale secure cloud infrastructure, deployments, observability, compliance, and incident response for a hardware collaboration platform. The role requires substantial cloud or security engineering experience, AWS and Linux expertise, and the ability to lead cross-functional infrastructure initiatives.