# Staff Infrastructure Engineer

**Company:** [Polymarket](https://hotfix.jobs/companies/polymarket)
**Location:** New York, NY
**Role:** DevOps / SRE
**Experience:** 8+ years
**Skills:** Kubernetes, Docker, Go, Python, AWS, GCP, Terraform, Pulumi, CI/CD, Linux, Prometheus, Grafana, Datadog, OpenTelemetry, Argo CD
**Posted:** 2026-08-25

> Build and operate Kubernetes-based infrastructure, cloud systems, CI/CD, observability, and reliability tooling for a high-scale prediction market platform. The role requires 8+ years of infrastructure or platform engineering experience and strong production expertise with Kubernetes, cloud providers, infrastructure as code, and software development.

## Job Description

## Responsibilities
- Design, build, and operate Polymarket’s Kubernetes-based platform, including cluster architecture, workload scheduling, and multi-environment management.
- Build systems for autoscaling, load balancing, service discovery, and disaster recovery.
- Write production-grade code and internal tooling to manage infrastructure as code and eliminate operational toil.
- Build and maintain reliable CI/CD pipelines.
- Develop monitoring, logging, and alerting systems; participate in on-call rotations and lead root-cause analysis for production incidents.
- Architect and optimize AWS or GCP infrastructure for cost, performance, and security.
- Collaborate with backend, exchange, and product engineering teams on platform performance and growth.

## Requirements
- 8+ years of professional infrastructure, platform, or DevOps engineering experience, ideally operating high-traffic production systems.
- Deep production experience with Kubernetes and Docker, including cluster operations, workload orchestration, and container lifecycle management.
- Strong software engineering fundamentals and ability to write production-grade code in Go, Python, or a similar language.
- Experience with AWS or GCP, including networking, IAM, compute, and storage.
- Experience with infrastructure-as-code tooling such as Terraform or Pulumi and CI/CD systems.
- Ability to debug complex distributed production systems on Linux and own correctness, performance, and uptime.

## Nice-to-haves
- Experience operating high-throughput, low-latency, or financial/trading infrastructure.
- Experience with service mesh, GitOps tools such as ArgoCD or Flux, or multi-cluster/multi-region architectures.
- Familiarity with observability stacks such as Prometheus, Grafana, Datadog, or OpenTelemetry.

## Compensation and Benefits
- Competitive salary and equity.
- Unlimited PTO.
- Full health, vision, and dental coverage.
- 401(k) match.
- Hardware setup including a new MacBook Pro, large display, and accessories.

## Similar jobs

- [Staff+ Site Reliability Engineer, Safeguards ML Infra](https://hotfix.jobs/jobs/6492550a-2ff4-498b-8247-470adae7d0c3) - Anthropic - San Francisco, CA - $320k – $485k/yr
- [Staff Infrastructure Engineer](https://hotfix.jobs/jobs/6e086905-4170-407a-a7a4-2e05df0701c5) - Polymarket - New York, NY - $250k – $500k/yr
- [Senior/Staff Infrastructure & Platform Engineer](https://hotfix.jobs/jobs/503b2675-718e-49a4-9692-0ca04a50b707) - Fortanix - Santa Clara, CA - $155k – $230k/yr
- [Staff Network Engineer, App Platform](https://hotfix.jobs/jobs/a410525c-f62d-4037-a6a5-40fa01a905e7) - Scale AI - San Francisco, CA
- [Staff Platform Engineer](https://hotfix.jobs/jobs/d1028610-5698-4ea7-850d-04183a6658da) - Motive - Buffalo, NY - $164k – $236k/yr

**Apply:** https://hotfix.jobs/jobs/024161a0-1250-463a-a4a6-a6da0cf75d74
**Canonical:** https://hotfix.jobs/jobs/024161a0-1250-463a-a4a6-a6da0cf75d74