Skip to content
Cerebras SystemsCerebras SystemsSunnyvale, CA

Infrastructure Engineer

Infrastructure Engineer responsible for hands-on installation, provisioning, maintenance, and troubleshooting of high-performance on-premise server hardware, Linux systems, and high-speed networking (100G/400G) in a data center environment. Requires 3+ years experience with Linux admin, x86 hardware, and network configuration.

Salary not listed
On-site3+ YOEDevOps / SRE

About the role

Key Responsibilities

  • Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)
  • Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link status
  • Configure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)
  • Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount points
  • Debug network issues at physical and OS level (VLAN, link issues, routing, etc.)
  • Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issues
  • Follow provisioning playbooks and maintain accurate records of assets and changes
  • Use scripting (Bash, Python) to automate routine tasks and improve efficiency
  • Collaborate with internal teams (network, systems, storage) and coordinate vendor RMAs
  • Document procedures and contribute to team knowledge base
  • Troubleshoot and replace failed server components with minimal downtime

Qualifications

  • 3–5+ years of experience in data center, lab, or infrastructure engineering roles
  • Proficient in Linux system administration and network configuration
  • Strong hands-on knowledge of x86 server hardware and enterprise networking
  • Familiar with BIOS configuration, firmware updates, and remote management tools
  • Skilled in physical setup and troubleshooting of high-speed NICs and optical links
  • Experience with VLANs, static routing, and diagnosing layer 1–3 issues
  • Ability to write scripts for automation and diagnostics (Bash, Python preferred)
  • Comfortable working on-site daily and lifting/moving server hardware

Preferred Skills

  • Experience with PXE, NFS, RAID controllers, and monitoring tools
  • Familiarity with configuration management tools (e.g., Ansible)
  • Prior experience in a lab or R&D hardware/software environment

Skills

Linuxdata center operationsserver hardwareNetworkingBashPythonipmiidracilovlanAnsiblepxenfsraid

Similar roles

DevOps / SRE jobs
Elicit

Infrastructure Engineer

ElicitOakland, CA

Own and evolve Elicit's cloud infrastructure platform (AWS/GCP, Kubernetes, Terraform) to support scalable single-tenant enterprise deployments. Build observability, compliance (SOC 2), cost optimization, and developer experience while contributing to backend systems where infra meets application logic. Requires 5+ years infrastructure/SRE experience, strong Terraform and K8s expertise, and enthusiasm for AI coding agents.

Salary not listed
On-site5+ YOEDevOps / SRE
OpenAI

Simulation Environments Engineer

OpenAISan Francisco, CA

Build and maintain CI/CD pipelines, orchestration, and automation for large-scale robotics simulation (SIL/HIL) to support model training, evaluation, and RL workloads at OpenAI. Requires strong infra, distributed systems, and Python/C++/Rust experience.

230k – 385k/yr
Hybrid5+ YOEDevOps / SRE
Airbnb

Operations Engineer, BizTech

AirbnbUnited States

Operations Engineer using AI, LLMs, and intelligent automation to triage tickets, accelerate incident response, build self-healing observability, and automate repetitive operational work in Airbnb's BizTech Global Operations team.

136k – 160k/yr
Remote3+ YOEDevOps / SRE
OpenAI

Systems Integration Engineer, Build Systems | Consumer Devices

OpenAISan Francisco, CA

Build and evolve Bazel, Yocto, and Buildkite-based CI systems for OpenAI consumer device software. Focus on hermetic builds, remote caching, test optimization, observability, and AI-powered failure analysis to accelerate reliable shipping. Requires 5+ years building developer infrastructure at scale.

293k – 325k/yr
Hybrid5+ YOEDevOps / SRE
Airbnb

Software Engineer, CI Platform Infrastructure

AirbnbUnited States

Build and optimize a next-generation CI platform infrastructure for workflow orchestration, scheduling, caching, and autoscaling to accelerate software development for engineers and AI coding agents at scale. Requires interest in distributed systems and knowledge of Kubernetes, EC2, Golang, and Docker.

162k – 190k/yr
RemoteDevOps / SRE