Infrastructure Engineer responsible for hands-on installation, provisioning, maintenance, and troubleshooting of high-performance on-premise server hardware, Linux systems, and high-speed networking (100G/400G) in a data center environment. Requires 3+ years experience with Linux admin, x86 hardware, and network configuration.
Salary not listed
On-site3+ YOEDevOps / SRE
About the role
Key Responsibilities
Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)
Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link status
Configure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)
Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount points
Debug network issues at physical and OS level (VLAN, link issues, routing, etc.)
Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issues
Follow provisioning playbooks and maintain accurate records of assets and changes
Use scripting (Bash, Python) to automate routine tasks and improve efficiency
Collaborate with internal teams (network, systems, storage) and coordinate vendor RMAs
Document procedures and contribute to team knowledge base
Troubleshoot and replace failed server components with minimal downtime
Qualifications
3–5+ years of experience in data center, lab, or infrastructure engineering roles
Proficient in Linux system administration and network configuration
Strong hands-on knowledge of x86 server hardware and enterprise networking
Familiar with BIOS configuration, firmware updates, and remote management tools
Skilled in physical setup and troubleshooting of high-speed NICs and optical links
Experience with VLANs, static routing, and diagnosing layer 1–3 issues
Ability to write scripts for automation and diagnostics (Bash, Python preferred)
Comfortable working on-site daily and lifting/moving server hardware
Preferred Skills
Experience with PXE, NFS, RAID controllers, and monitoring tools
Familiarity with configuration management tools (e.g., Ansible)
Prior experience in a lab or R&D hardware/software environment
Skills
Linuxdata center operationsserver hardwareNetworkingBashPythonipmiidracilovlanAnsiblepxenfsraid
Own and evolve Elicit's cloud infrastructure platform (AWS/GCP, Kubernetes, Terraform) to support scalable single-tenant enterprise deployments. Build observability, compliance (SOC 2), cost optimization, and developer experience while contributing to backend systems where infra meets application logic. Requires 5+ years infrastructure/SRE experience, strong Terraform and K8s expertise, and enthusiasm for AI coding agents.
Salary not listed
On-site5+ YOEDevOps / SRE
Simulation Environments Engineer
OpenAISan Francisco, CA
Build and maintain CI/CD pipelines, orchestration, and automation for large-scale robotics simulation (SIL/HIL) to support model training, evaluation, and RL workloads at OpenAI. Requires strong infra, distributed systems, and Python/C++/Rust experience.
230k – 385k/yr
Hybrid5+ YOEDevOps / SRE
Operations Engineer, BizTech
AirbnbUnited States
Operations Engineer using AI, LLMs, and intelligent automation to triage tickets, accelerate incident response, build self-healing observability, and automate repetitive operational work in Airbnb's BizTech Global Operations team.
136k – 160k/yr
Remote3+ YOEDevOps / SRE
Systems Integration Engineer, Build Systems | Consumer Devices
OpenAISan Francisco, CA
Build and evolve Bazel, Yocto, and Buildkite-based CI systems for OpenAI consumer device software. Focus on hermetic builds, remote caching, test optimization, observability, and AI-powered failure analysis to accelerate reliable shipping. Requires 5+ years building developer infrastructure at scale.
293k – 325k/yr
Hybrid5+ YOEDevOps / SRE
Software Engineer, CI Platform Infrastructure
AirbnbUnited States
Build and optimize a next-generation CI platform infrastructure for workflow orchestration, scheduling, caching, and autoscaling to accelerate software development for engineers and AI coding agents at scale. Requires interest in distributed systems and knowledge of Kubernetes, EC2, Golang, and Docker.