Leads strategy, reliability, capital planning, and operations for critical infrastructure across multiple data centers. The role requires 10+ years of facilities or critical infrastructure experience, including five years leading multi-site or multi-facility teams.
Salary not listed
On-site10+ YOEEngineering Management
About the role
Responsibilities
Lead and develop a multi-layered organization of managers, supervisors, and technicians across multiple data centers to deliver 24/7/365 reliability with zero unplanned interruptions.
Set site-wide strategy, multi-year roadmaps, CapEx/OpEx budgets, and resource forecasts for data center facilities infrastructure.
Drive execution, safety culture, and reliability engineering across the organization.
Partner with engineering, IT, operations, and executive leadership on large-scale projects supporting capacity expansion and efficiency improvements.
Establish standardized work instructions, maintenance strategies, and operational processes across data centers.
Assess resource needs, lifecycle costs, and capacity for multiple facilities; own capital planning and total cost of ownership.
Oversee contractors and vendors supporting mechanical, structural, HVAC, electrical, power, cooling, fire/life safety, and water systems.
Own preventive and predictive maintenance programs, including strategies for repairs, upgrades, replacements, uptime, efficiency, and PUE optimization.
Define and communicate KPIs covering safety, uptime, quality, labor productivity, energy efficiency, and operating expenses.
Provide strategic oversight of emergency response and on-call frameworks, including 24/7 support and rapid recovery capabilities.
Lead talent development, succession planning, and organizational design for a high-performing facilities team.
Requirements
Bachelor’s degree in engineering, facilities management, architecture, construction management, or a related technical field.
10+ years of progressive experience in facilities operations, data center infrastructure, or critical facilities management.
At least 5 years of leadership experience managing multi-site or multi-facility operations.
Experience leading teams responsible for high-availability environments such as data centers, mission-critical facilities, or industrial facilities.
Willingness to travel as needed across the site and occasionally to other company locations.
Ability and willingness to work extended hours, weekends, and non-standard schedules to support critical operations.
Ability to operate in a high-stress, high-tempo environment while balancing multiple strategic priorities.
Nice-to-Haves
Master’s degree in engineering, business, or a related field.
Professional certifications such as CEM, PMP, CFM, or data center-specific credentials.
Expertise in high-power electrical systems, UPS and generators, precision cooling/HVAC, fire suppression, BMS/DCIM systems, telecommunications infrastructure, fuel systems, and redundancy/resiliency design.
Experience with Tier standards and concurrent maintainability.
Track record managing large budgets, multi-year capital programs, and complex vendor ecosystems across multiple facilities.
Excellent executive communication and technical-operational presentation skills.
Strong leadership presence and experience building high-performing teams and driving cultural change.
Experience contributing to facility design, capacity planning, and long-term infrastructure roadmaps.
Experience managing large-scale projects and portfolios under aggressive timelines and changing priorities.
Skills
data center infrastructurehvacelectrical systemsupsgeneratorsfire suppressionbms/dcimpreventive maintenancepredictive maintenanceCapacity Planningpue optimizationcapital planningVendor Managementreliability engineeringpower systems
Hands-on technical leader owning end-to-end engineering for AI agents in compliance workflows, scaling team from 5 to 20+ engineers while architecting production-grade LLM systems and driving weekly shipping velocity. Requires founder experience, AI expertise, and Bay Area network.
Salary not listedOn-siteEngineering Management
Director of Engineering, Core & Ads Serving Platform
PinterestSan Francisco, CA +1
Lead technical strategy and execution for Pinterest's indexing and retrieval infrastructure across Core, Ads, and Shopping. Drive modernization of real-time and incremental systems to improve freshness, quality, relevance, and cost efficiency at massive scale while mentoring senior engineers.
285k – 450k/yrHybrid8+ YOEEngineering Management
Director, Analytics Systems
Beacon BiosignalsBoston, MA
Lead and grow a team of 6-10 engineers building Beacon's Analytics Systems for clinical data curation, biomarker computation, and integration of foundation models. Requires 6+ years managing analytics engineering teams in regulated biotech or medical device environments.
Salary not listedHybrid8+ YOEEngineering Management
Director, Development
CrusoeSan Francisco, CA
Leads execution of multi-billion dollar AI data center campuses, overseeing planning, construction, permitting, contractor management, and risk mitigation. Requires 10+ years in real estate development, technical project oversight, and municipal coordination skills.
Lead the strategy, architecture, and hands-on implementation of an end-to-end AI-native full-stack platform as a senior executive. Requires deep expertise in AI/ML and full-stack development, team leadership, strategic vision, and startup experience to bridge technical and business needs.
Salary not listedOn-site8+ YOEEngineering Management