Supervisor, Data Center Operations
Supervise data center technicians while overseeing server and network infrastructure installation, maintenance, troubleshooting, and operational improvement. The role requires 5+ years of relevant hardware and repair experience, technical leadership, Linux proficiency, and scripting experience.
About the job
Responsibilities
- Lead and mentor a team of data center technicians, fostering a culture of excellence and continuous improvement.
- Oversee the installation, maintenance, and troubleshooting of server and network infrastructure.
- Manage and optimize data center operations, including power supply cabling, fiber/optics labeling, and hardware decommissioning.
- Develop and enforce standard operating procedures (SOPs) and ensure adherence to safety protocols.
- Coordinate with engineering and provisioning teams to ensure seamless hardware intake and repair processes.
- Utilize internal applications for inventory and asset management, ensuring accurate tracking and reporting.
- Manage data center operations tickets via Jira, ensuring timely resolution and documentation.
- Collaborate with cross-functional teams to design and implement network layouts and solutions.
- Lead initiatives to improve operational efficiency and reduce downtime.
- Provide on-call support and respond to critical events as needed.
Requirements
- High school diploma or equivalency certificate.
- 5+ years of experience working with server, storage, compute, and network hardware.
- 5+ years of experience troubleshooting and repairing servers and networking infrastructure.
- Proven leadership experience in a data center or technical operations environment.
- Strong Linux skills, including navigating system directories, manipulating files in the Linux shell, configuring user permissions, and installing packages.
- Experience with Python, Bash, or other scripting languages.
- Experience leading data center infrastructure projects.
- Familiarity with structured cabling, including copper and fiber, and power and cooling concepts inside the data center.
- Excellent prioritization and time management skills.
- Ability to work in a fast-paced environment and maintain attention to detail.
Additional Requirements
- Ability to lift up to 35 lbs. unassisted.
- Comfortable working at elevated heights up to 50 feet with appropriate safety gear.
- Comfortable working in an environment requiring exposure to noise.
- Available to work evenings and weekends as schedules vary according to site operational needs.
Compensation and Benefits
- Position is subject to pre-employment and annual post-employment background checks.
Skills
Linux, Python, Bash, Jira, Server Hardware, Network Hardware, Structured Cabling, Fiber Optics, Power And Cooling, Asset Management
Similar jobs
DevOps / SRE jobsBuild and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.
Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.