Skip to content
HiveHive

Data Center Technician

Maintains and expands data centers supporting AI/ML infrastructure by installing, troubleshooting, and repairing servers, networks, and hardware. Requires 3-5 years data center experience, Linux knowledge, physical lifting ability, and on-call availability.

About the job

Responsibilities

  • Install and upgrade data center equipment racks, including switches, routers, monitoring systems and other large scale networking gear
  • Monitor, audit, and perform ongoing diagnostics, maintenance, and/or decommissioning on existing and new data center servers and network infrastructure
  • Maintain integration and deployment tooling
  • Participate in on-call rotation and root cause analysis as needed to respond to server, network, and hardware issues
  • Maintain an inventory and event logs of data center processes
  • Complete assigned tickets to uphold tight SLAs and respond to requests in a timely manner
  • Report actual or suspected security and/or policy violations/breaches to an appropriate authority

Requirements

  • Associate’s Degree or equivalent practical experience and knowledge of various Linux Distros
  • Minimum 3-5 years of experience in data center environments, operations, IT, or related fields
  • Ability to work long and/or on-call shifts that may include evenings/nighttimes, weekends, and/or holidays
  • Ability to physically lift equipment at least 50 pounds
  • Exceptional attention to detail and ability to troubleshoot intricate terminal systems, swap out failed components, and repair servers (discrete/rack-based)
  • Knowledge of the installation of software and firmware updates, OS, PXE
  • Familiarity with RAID, SAN, x86 architecture, Command Line Interface, Boot Processes, GRUB/LILO, File Systems, network device and protocol configuration
  • RMA processing and coordination with the Logistics Team
  • Handle storage media/Data Bearing Device (DBD) Reconcile/Physical Audits
  • Flexibility to travel to various data centers as needed

Preferred Qualifications

  • Applicable certifications: CompTIA (Server+, Network+) or CCNP
  • Networking: TCP/IP, ICMP, SSH, DNS, HTTP, SSL/TLS, Storage systems, RAID, distributed file systems, NFS/iSCSI/CIFS
  • Core OS Services: SSH, telnet, FTP, NFS, DNS, DHCP, LDAP
  • Experience with NVIDIA GPU linux software stack
  • Configuration Management - Chef
  • Version Control - GitHub
  • Containerization - Docker
  • Container Orchestrators - Mesosphere/Kubernetes
  • Virtualization - QEMU/KVM
  • Network hardware - Arista/Cisco/Fortinet

Skills

Linux, Raid, San, X86 Architecture, Command Line Interface, Pxe, Nvidia Gpu, Docker, Kubernetes, Chef, GitHub, Cisco, Arista, Nfs, Ssh

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.