# Senior Site Reliability Engineer, Kong Konnect

**Company:** [Kong](https://hotfix.jobs/companies/kong)
**Location:** Olympia, WA
**Role:** DevOps / SRE
**Salary:** $113k – $162k/yr
**Experience:** 5+ years
**Skills:** Kubernetes, Terraform, terragrunt, Helm, Argo CD, Postgres, Redis, ClickHouse, druid, kong gateway, kong mesh, Prometheus, Grafana, AWS, Kafka
**Posted:** 2025-11-05

> Senior Site Reliability Engineer responsible for operating and scaling Kong’s multi-region SaaS platform across AWS, GCP, and Azure. The role focuses on Kubernetes infrastructure, automation, observability, reliability engineering, and production operations at enterprise scale.

## Job Description

## Responsibilities
- Operate and scale Kong’s global SaaS platform, Konnect, across regions and cloud providers.
- Build, automate, and maintain Kubernetes infrastructure and deployment workflows using Terraform, Terragrunt, Helm, and ArgoCD.
- Design, maintain, and optimize multi-region data and caching layers, including PostgreSQL, Redis, ClickHouse, and Druid.
- Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.
- Develop and maintain CI/CD pipelines and GitOps workflows.
- Improve observability and incident-response readiness using Datadog, Prometheus, Grafana, and Thanos; define and track SLOs.
- Collaborate with development and security teams to operate SaaS services in compliance with reliability, security, and regulatory standards.
- Participate in a global 24/7 on-call rotation and improve operational playbooks and postmortem practices.
- Lead scaling initiatives that improve elasticity, reliability, and cost efficiency.

## Requirements
- Bachelor’s degree in Computer Science or equivalent practical experience.
- Experience managing enterprise-scale SaaS or PaaS systems in multi-region, multi-tenant, secure environments.
- Deep Kubernetes expertise, including cluster and networking troubleshooting and fault-tolerant, scalable design.
- Strong proficiency with infrastructure-as-code tools such as Terraform or Terragrunt.
- Experience with CI/CD pipelines and GitOps workflows, including ArgoCD, Atlantis, and Helm.
- Proficiency in Go, Python, or Bash for automation and tooling.
- Strong understanding of Linux/Unix systems, DNS, TLS/SSL, HTTP, load balancers, and distributed systems.
- Experience with API gateway and service mesh technologies.
- Familiarity with Kafka and observability platforms such as Datadog, Prometheus, and Grafana.
- Experience working in a 24/7/365 production-support environment.

## Nice-to-Haves
- Hands-on experience with Kong Gateway, Kong Mesh, or similar service-connectivity technologies.
- Experience operating ClickHouse, Druid, or other time-series and analytics databases.
- Experience managing PostgreSQL and Redis in multi-region configurations.
- Working knowledge of AWS networking, Azure VNet, or GCP NCC.
- Strong understanding of disaster recovery, resiliency testing, and compliance-driven reliability practices.

## Similar roles

- [Senior Network Engineer](https://hotfix.jobs/jobs/c7465a12-34cf-4779-af27-910c7aea361e) - Black Canyon Consulting - Bethesda, MD - $110k – $160k/yr
- [Senior Network Engineer (Py)](https://hotfix.jobs/jobs/931ade13-093a-4f63-b296-7e1a816b7786) - Black Canyon Consulting - Bethesda, MD - $110k – $160k/yr
- [Senior Engineer, Autonomy Systems Integration and Test (R4472)](https://hotfix.jobs/jobs/dba713b9-cc19-4d0c-ae65-8473dc698966) - Shield AI - San Diego, CA - $110k – $170k/yr
- [Senior SRE, Managed Gateways](https://hotfix.jobs/jobs/ad32fcba-eefa-4554-820d-95dba7cbf225) - Kong - Remote - $118k – $167k/yr
- [Senior Engineer, Platform Infrastructure](https://hotfix.jobs/jobs/cf15e8c3-ee2a-4f6a-abd6-8a54bce776fd) - Shield AI - San Diego, CA - $120k – $180k/yr

**Apply:** https://hotfix.jobs/jobs/34f08e92-3676-4944-9c37-7e5e58d24f70
**Canonical:** https://hotfix.jobs/jobs/34f08e92-3676-4944-9c37-7e5e58d24f70