Skip to content
KongKongOlympia, WA

Senior Site Reliability Engineer, Kong Konnect

Senior Site Reliability Engineer responsible for operating and scaling Kong’s multi-region SaaS platform across AWS, GCP, and Azure. The role focuses on Kubernetes infrastructure, automation, observability, reliability engineering, and production operations at enterprise scale.

113k – 162k/yr
On-site5+ YOEDevOps / SRE

About the role

Responsibilities

  • Operate and scale Kong’s global SaaS platform, Konnect, across regions and cloud providers.
  • Build, automate, and maintain Kubernetes infrastructure and deployment workflows using Terraform, Terragrunt, Helm, and ArgoCD.
  • Design, maintain, and optimize multi-region data and caching layers, including PostgreSQL, Redis, ClickHouse, and Druid.
  • Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.
  • Develop and maintain CI/CD pipelines and GitOps workflows.
  • Improve observability and incident-response readiness using Datadog, Prometheus, Grafana, and Thanos; define and track SLOs.
  • Collaborate with development and security teams to operate SaaS services in compliance with reliability, security, and regulatory standards.
  • Participate in a global 24/7 on-call rotation and improve operational playbooks and postmortem practices.
  • Lead scaling initiatives that improve elasticity, reliability, and cost efficiency.

Requirements

  • Bachelor’s degree in Computer Science or equivalent practical experience.
  • Experience managing enterprise-scale SaaS or PaaS systems in multi-region, multi-tenant, secure environments.
  • Deep Kubernetes expertise, including cluster and networking troubleshooting and fault-tolerant, scalable design.
  • Strong proficiency with infrastructure-as-code tools such as Terraform or Terragrunt.
  • Experience with CI/CD pipelines and GitOps workflows, including ArgoCD, Atlantis, and Helm.
  • Proficiency in Go, Python, or Bash for automation and tooling.
  • Strong understanding of Linux/Unix systems, DNS, TLS/SSL, HTTP, load balancers, and distributed systems.
  • Experience with API gateway and service mesh technologies.
  • Familiarity with Kafka and observability platforms such as Datadog, Prometheus, and Grafana.
  • Experience working in a 24/7/365 production-support environment.

Nice-to-Haves

  • Hands-on experience with Kong Gateway, Kong Mesh, or similar service-connectivity technologies.
  • Experience operating ClickHouse, Druid, or other time-series and analytics databases.
  • Experience managing PostgreSQL and Redis in multi-region configurations.
  • Working knowledge of AWS networking, Azure VNet, or GCP NCC.
  • Strong understanding of disaster recovery, resiliency testing, and compliance-driven reliability practices.

Skills

KubernetesTerraformterragruntHelmArgo CDPostgresRedisClickHousedruidkong gatewaykong meshPrometheusGrafanaAWSKafka

Similar roles

DevOps / SRE jobs
Shield AI

Senior Engineer, Network Integration (R4926)

Shield AIDallas, TX

Design, implement, and optimize enterprise network infrastructures including LAN, WAN, cloud, and hybrid environments with focus on security, high availability, and data center infrastructure. Requires 8+ years of network engineering experience and deep expertise in cybersecurity protocols.

110k – 170k/yrOn-site8+ YOEDevOps / SRE
Black Canyon Consulting

Senior Network Engineer

Black Canyon ConsultingBethesda, MD

Leads design, implementation, and automation of high-performance networks using Arista EOS, CloudVision, F5 BIG-IP, and protocols like BGP, VxLAN. Requires 10+ years experience, Python scripting, and team leadership for NIH scientific mission.

110k – 160k/yrHybrid10+ YOEDevOps / SRE
Black Canyon Consulting

Senior Network Engineer (Py)

Black Canyon ConsultingBethesda, MD

Designs and deploys high-performance network architectures using Spine-and-Leaf, Arista EOS/CloudVision, and F5 BIG-IP. Automates operations with Python, modernizes legacy networks for cloud, and leads technical efforts requiring 10+ years experience and expertise in BGP, OSPF, VxLAN, EVPN.

110k – 160k/yrHybrid10+ YOEDevOps / SRE
Shield AI

Senior Engineer, Autonomy Systems Integration and Test (R4472)

Shield AISan Diego, CA

Leads hands-on integration and testing of autonomy systems on small unmanned aircraft, debugging issues across hardware, software, and flight environments. Requires 4-7 years experience in robotics/aerospace testing, Python/C++ proficiency, and bachelor's in engineering or related field.

110k – 170k/yrOn-site4+ YOEDevOps / SRE
Kong

Senior SRE, Managed Gateways

KongOlympia, WA

Senior Site Reliability Engineer owning production reliability and enterprise customer implementations for Kong's fast-growing Managed Gateways SaaS product across AWS, GCP, and Azure. Requires deep Kubernetes, cloud-native, and Golang expertise plus customer-facing technical leadership.

118k – 167k/yrRemote7+ YOEDevOps / SRE