Skip to content

Infrastructure Engineer

Infrastructure Engineer responsible for hands-on installation, provisioning, maintenance, and troubleshooting of high-performance on-premise server hardware, Linux systems, and high-speed networking (100G/400G) in a data center environment. Requires 3+ years experience with Linux admin, x86 hardware, and network configuration.

About the job

Key Responsibilities

  • Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)
  • Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link status
  • Configure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)
  • Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount points
  • Debug network issues at physical and OS level (VLAN, link issues, routing, etc.)
  • Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issues
  • Follow provisioning playbooks and maintain accurate records of assets and changes
  • Use scripting (Bash, Python) to automate routine tasks and improve efficiency
  • Collaborate with internal teams (network, systems, storage) and coordinate vendor RMAs
  • Document procedures and contribute to team knowledge base
  • Troubleshoot and replace failed server components with minimal downtime

Qualifications

  • 3–5+ years of experience in data center, lab, or infrastructure engineering roles
  • Proficient in Linux system administration and network configuration
  • Strong hands-on knowledge of x86 server hardware and enterprise networking
  • Familiar with BIOS configuration, firmware updates, and remote management tools
  • Skilled in physical setup and troubleshooting of high-speed NICs and optical links
  • Experience with VLANs, static routing, and diagnosing layer 1–3 issues
  • Ability to write scripts for automation and diagnostics (Bash, Python preferred)
  • Comfortable working on-site daily and lifting/moving server hardware

Preferred Skills

  • Experience with PXE, NFS, RAID controllers, and monitoring tools
  • Familiarity with configuration management tools (e.g., Ansible)
  • Prior experience in a lab or R&D hardware/software environment

Skills

Linux, Data Center Operations, Server Hardware, Networking, Bash, Python, Ipmi, Idrac, Ilo, Vlan, Ansible, Pxe, Nfs, Raid

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.