Deployment Manager
Lead end-to-end deployment of large-scale AI clusters in data centers, from facility readiness to customer handoff. Manage rack integration, high-density cabling, facilities coordination, troubleshooting, and cross-functional execution for hyperscale AI infrastructure.
About the job
Responsibilities
- Lead deployment of AI clusters including Cerebras Wafer Scale Engine, high-speed switches, storage, and rack-level infrastructure.
- Coordinate rack integration of high-density compute, specialized racks, and associated power components (PDUs, busway drops).
- Ensure deployment aligns with AI cluster architecture, topology, and scaling requirements.
- Manage installation of high-speed interconnect cabling (fiber and copper) supporting AI fabrics (east–west traffic) in Data Centers.
- Coordinate inter-rack and intra-rack cabling for AI clusters, including spine-leaf and pod-level designs.
- Ensure proper routing, airflow clearance, labeling, and testing of all AI-related cabling.
- Work closely with facilities teams on power capacity, cooling readiness, containment, and grounding for dense racks.
- Coordinate deployment sequencing with facility commissioning milestones.
- Validate white-space readiness before rack and cluster deployment.
- Provide directions to technicians on the data center floor.
- Troubleshoot cabling, connectivity, and integration issues impacting AI cluster bring-up.
- Lead root-cause analysis for deployment blockers related to cabling, hardware placement, or facilities dependencies.
- Support validation, burn-in, and handoff of AI clusters to operations teams.
- Partner with network, server, AI platform, and operations teams to align on deployment plans and readiness.
- Manage multiple parallel AI cluster deployments across sites or availability zones.
- Communicate risks, dependencies, and milestones clearly to stakeholders.
- Ensure deployments follow company design standards, structured cabling best practices, and AI deployment playbooks.
- Validate as-built documentation, labeling accuracy, and deployment checklists.
- Maintain accurate records for cluster configuration, cabling, and deployment status.
- Enforce EHS, data center safety, and access control procedures during deployment.
- Ensure safe handling of heavy, high-power GPU equipment.
- Proactively identify and mitigate deployment and operational risks.
Qualifications
- Bachelor’s degree in Engineering, IT, or equivalent practical experience.
- 10+ years of experience in data center deployments, infrastructure delivery, or integration roles.
- Ability to provide directions to technicians on data center floor.
- Hands-on experience deploying hyperscale AI, ML, or HPC infrastructure.
- Strong experience with structured cabling in high-density environments, and troubleshooting.
- Proven ability to manage complex, cross-functional deployment programs.
- Familiarity with high-speed fabrics (e.g., InfiniBand, high-bandwidth Ethernet).
- Experience with DCIM or deployment tracking systems.
- Strong attention to detail and operational rigor.
- Ability to perform under tight timelines and production constraints.
Nice-to-Haves
- Experience with high-power, high-thermal-density environments.
- Familiarity with AI cluster architecture, topology, and scaling.
Skills
Data Center Deployment, Rack Integration, High-Density Cabling, Structured Cabling, InfiniBand, High-Bandwidth Ethernet, Dcim, AI Infrastructure, Hpc Infrastructure, Troubleshooting, Cross-Functional Coordination, Facilities Coordination
Similar jobs
Solutions Architecture jobsBuild and deploy Kong API and AI connectivity solutions for enterprise customers, leading migrations, automation, integrations, and production implementations across cloud-native environments. Requires 8+ years of engineering experience, strong Kubernetes and cloud expertise, and customer-facing technical leadership.
Leads technical solution design, demonstrations, and proof-of-value engagements for enterprise customers in partnership with sales. Requires 7+ years of enterprise solutions engineering experience, strong data-platform expertise, and willingness to travel at least 30%.
Leads strategic technical pursuits for enterprise customers, advising on architecture, AI deployments, and complex sales cycles while translating field insights into product and go-to-market strategy. Requires 15+ years of technical leadership experience, enterprise architecture expertise, and strong executive communication.
Serve as a strategic technical advisor to enterprise prospects and customers, guiding evaluations, architecture, implementation, and adoption of Temporal. The role requires distributed-systems expertise, application prototyping skills, cloud experience, strong technical communication, and close partnership with sales.
Own the architecture of an AI development ecosystem that supports data, model, simulation, and deployment workflows for autonomous capabilities. The role requires 10+ years of software architecture experience, strong AI/ML and data architecture expertise, and the ability to lead across complex engineering teams.