Latest DevOps / SRE jobs at Hive
Job results
Maintains and expands data centers supporting AI/ML infrastructure by installing, troubleshooting, and repairing servers, networks, and hardware. Requires 3-5 years data center experience, Linux knowledge, physical lifting ability, and on-call availability.
Senior Site Reliability Engineer automates operational processes, manages secure infrastructure across hybrid data centers and cloud, and improves workflows for ML/data teams. Requires 3-5 years experience with Linux, containers, and automation tools.
Systems Engineer/DevOps role focused on automating operations, managing hybrid on-prem data centers and AWS infrastructure, and ensuring reliability for ML/SaaS services. Requires 1-2 years experience with Linux, containers, and automation tools.