Skip to content
NominalNominal

Baremetal Infrastructure Engineer

Deploy and support Nominal's self-hosted platform in customer environments including air-gapped and regulated sites. Own Linux, Kubernetes, and bare-metal infrastructure reliability while partnering directly with customer IT and security teams.

About the job

What You'll Do

  • Serve as a technical expert for customer-hosted and air-gapped deployments
  • Travel onsite to support critical customer implementations when additional technical expertise is required
  • Troubleshoot Linux, networking, Kubernetes, and infrastructure issues in production environments
  • Partner directly with customer IT, infrastructure, and security teams to gather requirements and ensure successful deployments
  • Build and improve deployment tooling, automation, and operational processes
  • Improve reliability, observability, and maintainability across our self-hosted infrastructure
  • Develop deployment documentation, runbooks, and troubleshooting guides
  • Help shape Nominal's long-term strategy for self-hosted and regulated deployments

Requirements

  • Strong Linux systems expertise
  • Experience operating and troubleshooting production infrastructure
  • Background in DevOps, SRE, Platform Engineering, or Infrastructure Engineering
  • Ability to independently debug complex systems across multiple layers of the stack
  • Comfort working directly with customers and external stakeholders
  • Willingness to travel for customer deployments
  • Strong ownership mentality and ability to operate in ambiguous environments

Nice-to-Haves

  • Active security clearance or ability and willingness to obtain and maintain one
  • Experience with Kubernetes in production environments
  • GitOps workflows and tooling (Flux, Helm, Kustomize, etc.)
  • Infrastructure as Code experience
  • Datacenter operations experience
  • Hardware procurement, provisioning, or lifecycle management
  • Networking expertise (routing, switching, troubleshooting, performance tuning)
  • Linux kernel, networking, or performance optimization experience
  • Security-focused infrastructure experience including TLS and certificate management
  • Experience supporting air-gapped, classified, or highly regulated environments

Skills

  • Bare Metal & Provisioning: PXE, IPMI, BMC, iLO / iDRAC, MaaS, Tinkerbell, Metal provisioning workflows
  • Systems: Linux internals, Kernel troubleshooting, NUMA, RDMA, SR-IOV, eBPF
  • Networking: BGP, Routing, Low-latency networking, NIC offloading, DPDK
  • Storage: Ceph, NVMe, Distributed storage systems
  • Infrastructure: Kubernetes, GitOps, Datacenter automation, Rack provisioning, Hardware orchestration, On-prem infrastructure operations

Benefits

  • 100% coverage of medical, dental, and vision insurance
  • Unlimited PTO and sick leave
  • Free lunch, snacks, and coffee
  • Professional Development Stipend
  • In-office hardware lab with a $250 project stipend
  • Annual company retreat

Skills

Linux, Kubernetes, DevOps, SRE, Infrastructure As Code, GitOps, Networking, Bare Metal Provisioning, Ceph, Ebpf

Fusion Health

Fusion Health

Woodbridge, NJ

DevOps Engineer
$120k+/yrHybrid5+ YOEDevOps / SRE

Owns secure, scalable Azure infrastructure for healthcare applications, including cloud migrations, Terraform-based automation, CI/CD pipelines, monitoring, and compliance. Requires 3–5+ years of Azure experience and strong DevOps and cloud-security expertise.

Kong

Kong

United States

Site Reliability Engineer 2
$123k+/yrRemoteDevOps / SRE

Operate and scale Kong’s multi-region SaaS platform across major cloud providers, Kubernetes, and distributed data systems. The role requires strong infrastructure automation, observability, CI/CD, and production reliability experience, with participation in a global on-call rotation.

PagerDuty

PagerDuty

Atlanta, GA

Site Reliability Engineer II
$113k+/yrHybrid3+ YOEDevOps / SRE

Operates and evolves foundational networking, compute, Kubernetes, and ingress infrastructure for PagerDuty’s real-time platform. Requires 3+ years in SRE, DevOps, or platform engineering, with Linux production operations, cloud infrastructure, programming, and Infrastructure as Code experience.

Mercor

Mercor

San Francisco, CA
Infrastructure Engineer
$130k+/yrOn-siteDevOps / SRE

Builds and scales highly available infrastructure using AWS, Terraform, and Docker to support rapid growth and AI workloads. Collaborates with product and research teams on architectures, CI/CD, monitoring, and performance optimization.

Mercor

Mercor

San Francisco, CA

Member of Technical Staff, Mercor Enterprise Platform
$130k+/yrOn-site5+ YOEDevOps / SRE

Build and operate Mercor’s enterprise agent platform across security, routing, isolated execution, orchestration, deployment, and production scalability. The role requires 5+ years building high-scale platforms, architectural ownership, and experience with core infrastructure primitives across multiple clouds.