Skip to content
Fab2Fab2

Infrastructure Engineer

Designs, deploys, and manages bare-metal on-prem infrastructure including servers, networking, observability, and backend services for a semiconductor fab. Requires hands-on SRE experience, systems programming in Rust/Go/Python, and BS in CS/CE or equivalent.

About the job

Responsibilities

  • Design and implement light-weight, performant, and reliable software infrastructure to power a semiconductor fab.
  • Procure, deploy and manage our fleet of on-prem servers, virtual machines, and single-board computers running on semiconductor fabrication equipment.
  • Deploy and manage backend services, e.g., consul, vault, grafana, victoria-metrics, alertmanager, redpanda, vector, gitea, postgres.
  • Design and setup low level networking components, e.g., service discovery, DNS, reverse proxies, TLS, S3 compatible storage, VPNs.
  • Scale our observability platform: Build systems to ingest and display both traditional system metrics as well as high frequency telemetry from semiconductor fabrication equipment.
  • Design and implement cross-site networking, replication and backups.
  • Automate OS image creation and deployment and software build and deployment systems.
  • Develop best practices and tools for security, authentication, authorization, and secrets management.
  • Help build our in-house infrastructure-as-code tool.

Required Experience

  • BS in Computer Science, Computer Engineering, or equivalent demonstrated exceptional skill in software engineering.
  • Demonstrated hands-on experience in backend infrastructure or Site Reliability Engineering, successfully and autonomously taking complex systems from concept to production.
  • Deep proficiency in at least one statically-typed, compiled language, with a strong command of systems fundamentals.

Nice-to-haves

  • Hands-on programming experience in Rust or Go.
  • Direct experience designing and managing on-premise, bare-metal infrastructure without relying on heavy cloud abstractions.

Skills

Linux, Systemd, Rust, Go, Python, Consul, Vault, Grafana, Victoriametrics, Alertmanager, Redpanda, Postgres, DNS, Tls, Vpn

Kong

Kong

United States

Site Reliability Engineer 2
$123k+/yrRemoteDevOps / SRE

Operate and scale Kong’s multi-region SaaS platform across major cloud providers, Kubernetes, and distributed data systems. The role requires strong infrastructure automation, observability, CI/CD, and production reliability experience, with participation in a global on-call rotation.

Mercor

Mercor

San Francisco, CA
Infrastructure Engineer
$130k+/yrOn-siteDevOps / SRE

Builds and scales highly available infrastructure using AWS, Terraform, and Docker to support rapid growth and AI workloads. Collaborates with product and research teams on architectures, CI/CD, monitoring, and performance optimization.

Mercor

Mercor

San Francisco, CA

Member of Technical Staff, Mercor Enterprise Platform
$130k+/yrOn-site5+ YOEDevOps / SRE

Build and operate Mercor’s enterprise agent platform across security, routing, isolated execution, orchestration, deployment, and production scalability. The role requires 5+ years building high-scale platforms, architectural ownership, and experience with core infrastructure primitives across multiple clouds.

Fusion Health

Fusion Health

Woodbridge, NJ

DevOps Engineer
$120k+/yrHybrid5+ YOEDevOps / SRE

Owns secure, scalable Azure infrastructure for healthcare applications, including cloud migrations, Terraform-based automation, CI/CD pipelines, monitoring, and compliance. Requires 3–5+ years of Azure experience and strong DevOps and cloud-security expertise.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.