# Director, Site Reliability Engineering

**Company:** [Okta](https://hotfix.jobs/companies/okta)
**Location:** Bengaluru, India
**Role:** Engineering Management
**Experience:** 16+ years
**Skills:** AWS, Kubernetes, Terraform, CI/CD, Infrastructure As Code, Observability, Automation, Incident Response, Performance Optimization, Cloud-Native Architecture, SaaS, Multi-Cloud, Root-Cause Analysis, Containerization, Finops
**Posted:** 2026-08-25

> Leads India-based Site Reliability Engineering teams responsible for Okta’s cloud platform, databases, networking, Kubernetes, CI/CD, observability, FinOps, and automation. Requires 16+ years in infrastructure or SRE, substantial people-management experience, and expertise in AWS, Kubernetes, Terraform, and reliable SaaS operations.

## Job Description

## Responsibilities
- Build and lead a high-caliber India-based SRE organization supporting Okta’s production fleet.
- Partner with global engineering, product, and infrastructure leaders to deliver resilient, scalable, and secure services.
- Define and execute the India SRE strategy in alignment with global reliability goals.
- Lead post-incident reviews, root-cause analysis, incident management, on-call rotations, and blameless RCAs.
- Implement automation and observability to reduce manual toil and improve operational efficiency.
- Drive adoption of infrastructure as code, container orchestration, and AI within infrastructure operations.
- Hire, mentor, and develop SRE talent across India while building a culture centered on reliability and innovation.
- Foster collaboration across U.S. and EMEA teams and promote continuous learning and process improvement.
- Manage service and business expectations and prioritize resource allocation.
- Develop robust platforms, tooling, and self-service capabilities to accelerate SRE and product engineering.
- Improve cloud infrastructure SDLC processes, including CI/CD, change management, and release management.

## Requirements
- 16+ years of experience in site reliability, infrastructure, or production engineering.
- 8+ years of technical leadership and people management experience, including managing managers.
- Experience building or scaling offshore SRE teams that partner with global counterparts.
- 4+ years running an SRE organization supporting a SaaS or cloud service in a public cloud, preferably AWS.
- Strong expertise in automation, observability, performance optimization, and incident response.
- Expertise in cloud-native architectures, Kubernetes, Terraform, and CI/CD pipelines.
- Demonstrated ability to lead cross-functional teams and manage large-scale programs.
- Excellent verbal, written, communication, interpersonal, and cross-cultural collaboration skills.
- Computer Science or related degree, or equivalent experience.

## Nice-to-haves
- Experience supporting a multi-cloud environment.
- Experience with AI applied to infrastructure operations.

## Compensation and Benefits
- In-person onboarding experience.
- Opportunities for developing talent, fostering connection and community, and supporting social impact.

## Similar jobs

- [Head of Engineering, AI Platform](https://hotfix.jobs/jobs/77d26cb0-f792-4400-9cce-e4fbe27da498) - Starburst
- [Director of Engineering, Organizations & Cells](https://hotfix.jobs/jobs/2ae2198e-46b2-41eb-bd6f-ce5b75caac9d) - GitLab - Remote
- [Director, Solutions Architecture](https://hotfix.jobs/jobs/f4e8a36e-29a5-4927-8ab3-2ef96e2e1a9d) - MongoDB - Remote
- [Director of Engineering](https://hotfix.jobs/jobs/2bc057ad-7f57-42ec-adbb-9a4a64a75a5f) - Databricks - Bengaluru, India
- [Director of Engineering](https://hotfix.jobs/jobs/04d05dcd-f763-43f2-b277-4c6139edab75) - Databricks - Bengaluru, India

**Apply:** https://hotfix.jobs/jobs/6be1147d-b8bf-49ce-99d9-bfb5a6fd78d1
**Canonical:** https://hotfix.jobs/jobs/6be1147d-b8bf-49ce-99d9-bfb5a6fd78d1