Skip to content
OktaOkta

SRE Operations Engineer

Supports the reliability and day-to-day operation of Okta’s Customer Identity Cloud by monitoring platform health, handling service requests, executing runbooks, and troubleshooting production issues. Requires cloud operations experience, infrastructure knowledge, and familiarity with Kubernetes and monitoring tools.

About the job

Responsibilities

  • Execute operational work, including updating, patching, and maintaining the Engineering Service Desk queue.
  • Triage and action team requests in a timely manner.
  • Monitor platform health and address deployment and operational issues.
  • Assist with capacity, performance, and scalability testing.
  • Serve as an escalation point for platform issues raised by customer support teams.
  • Execute runbooks and update operational processes as needed.
  • Interface with the SRE team to report core issues, improvement opportunities, and feature requests.

Requirements

  • General platform infrastructure knowledge, including high availability, load balancing, routers, firewalls, and storage subsystems.
  • Understanding of HTTP, SSL, SSH, and Kubernetes.
  • Familiarity with open-source technologies and tools such as MongoDB and Node.js.
  • Experience with monitoring and troubleshooting techniques.
  • Clear communication with diverse stakeholders across multiple domains.
  • Strong multitasking and time-management skills.
  • At least 1 year of experience in a cloud operations role.
  • At least 1 year of experience supporting large-scale, mission-critical applications in a production environment.
  • Interest in or understanding of programming, such as Go or shell scripting.

Nice to Have

  • Terraform knowledge.
  • Familiarity with AWS and Azure.
  • Linux fundamentals.
  • Experience with Datadog.

Skills

Kubernetes, Http, Ssl, Ssh, MongoDB, Node.js, Terraform, AWS, Azure, Linux, Datadog, Go, Shell Scripting, Load Balancing

Socure

Socure

Bengaluru, India

Software Engineer-II SRE
No salary listedOn-site2+ YOEDevOps / SRE

Build and operate reliable, scalable production systems across AWS, Kubernetes, infrastructure automation, CI/CD, and observability. The role requires 2–4 years of SRE, DevOps, platform, or cloud infrastructure experience and strong automation skills.

Kong

Kong

Bengaluru, India

Site Reliability Engineer 2, Managed Gateways
No salary listedOn-site2+ YOEDevOps / SRE

Owns reliability, scalability, and performance for managed gateway services by automating cloud operations, monitoring production systems, and resolving incidents. Requires at least two years of production SRE experience plus proficiency in Golang or Python, Kubernetes, and major cloud platforms.

StarTree

StarTree

India

Software Engineer SRE
No salary listedOn-site1+ YOEDevOps / SRE

Supports reliable, secure, and scalable cloud platforms across AWS, GCP, and Azure, with a focus on Kubernetes workloads. The role monitors services, troubleshoots incidents, supports deployments, and automates operations while requiring 1–2 years of SRE, DevOps, cloud operations, or infrastructure experience.

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Acryldata

Acryldata

Bengaluru, India

DevOps
No salary listedRemote5+ YOEDevOps / SRE

Own reliability, scalability, and operational excellence for DataHub Cloud and enterprise deployment offerings. The role requires 5+ years in DevOps, platform engineering, or SRE, with expertise in cloud platforms, Kubernetes, infrastructure as code, observability, and deployment automation.