# Staff Software Engineer - Infrastructure/DevOps

**Company:** [6sense](https://hotfix.jobs/companies/6sense)
**Location:** Bengaluru, India
**Role:** DevOps / SRE
**Experience:** 10+ years
**Skills:** Amazon Web Services, Kubernetes, Terraform, Pulumi, Ansible, Hashicorp Vault, Open Policy Agent, GitHub Actions, Jenkins, Prometheus, Grafana, Elk Stack, Datadog, Python, Go
**Posted:** 2026-08-13

> The role leads the design, automation, security, and reliability of multi-region cloud infrastructure and Kubernetes platforms. It requires extensive software and infrastructure engineering experience, strong AWS and infrastructure-as-code expertise, and proficiency in Python, Go, or Bash.

## Job Description

## Responsibilities
- Architect, build, and scale core infrastructure systems.
- Automate infrastructure lifecycle to minimize human intervention.
- Design secure connectivity across services, VPCs, accounts, and regions.
- Debug and resolve complex production issues across the stack.
- Write production-quality code, tools, and frameworks.
- Collaborate with engineering teams to standardize infrastructure usage.

## Core Infrastructure and Cloud Platform
- Design and evolve infrastructure on **Amazon Web Services** and **Kubernetes**.
- Build and scale multi-region and multi-account architectures.
- Implement secure and scalable service connectivity, including VPC, cross-account, and cross-region connectivity.

## Security and Zero Trust
- Drive identity-based access and eliminate shared credentials.
- Implement secrets management with **HashiCorp Vault**.
- Enforce policies with **Open Policy Agent**.
- Design secure service-to-service communication and fine-grained IAM and access control systems.

## Infrastructure as Code and Automation
- Build infrastructure using **Terraform**, **Pulumi**, and **Ansible**.
- Create reusable modules, abstractions, and fully automated provisioning workflows.
- Integrate infrastructure automation with **GitHub Actions** and **Jenkins**.

## Cluster and Infrastructure Lifecycle
- Own the lifecycle of Kubernetes clusters, including EKS, and databases such as RDS, Aurora, and ElastiCache.
- Automate cluster upgrades, migrations, node scaling, and patching.

## Multi-Region and Migration Engineering
- Design region-agnostic infrastructure patterns.
- Enable automated region bootstrapping and failover.
- Lead region, account, and cloud migration initiatives.

## Observability and Reliability
- Build and operate metrics, logging, and tracing systems using **Prometheus**, **Grafana**, the **ELK Stack**, and **Datadog**.
- Improve system reliability, alerting, and incident response.

## Governance and Standards
- Define and enforce resource-tagging strategies and data-classification policies.
- Ensure cost visibility and auditability.

## Developer Productivity and Platform
- Build self-service infrastructure workflows.
- Improve developer experience through automation and tooling.
- Contribute to internal platform evolution.

## Requirements
- 10+ years of experience in software engineering or equivalent experience.
- 5+ years of experience in infrastructure engineering.
- Strong hands-on experience with **Amazon Web Services**, **Kubernetes**, and **Terraform** or **Pulumi**.
- Strong understanding of networking, including VPC, routing, DNS, and connectivity.
- Strong understanding of IAM and access control.
- Experience automating infrastructure workflows.
- Experience with highly available and scalable systems.
- Proficiency in **Python**, **Go**, or **Bash**.

## Nice-to-Haves
- Multi-region or multi-account architecture experience.
- Experience automating infrastructure migrations and upgrades, including EKS upgrades.
- Experience with secrets-management platforms such as HashiCorp Vault.
- Familiarity with policy-as-code tools such as Open Policy Agent.
- Familiarity with observability systems such as Prometheus, Grafana, and Datadog.
- Exposure to data infrastructure such as Hadoop, Trino, and Spark.
- Exposure to service meshes such as Istio.

## Compensation and Benefits
- Health coverage.
- Paid parental leave.
- Paid time off and holidays.
- Quarterly self-care days off.
- Stock options.
- Equipment and support for working from home or in an office.
- Learning and development initiatives, including LinkedIn Learning.
- Wellness education sessions and employee resource group events.

## Similar jobs

- [Staff Software Engineer, Inference / Compute Infrastructure Engineering](https://hotfix.jobs/jobs/b1b1bdb1-d3f0-47cd-b94b-7781be7399fc) - Together AI - Remote
- [Staff DevSecOps Engineer, Enterprise Technology](https://hotfix.jobs/jobs/85d2a18a-f6f6-4858-a074-0d6b16869dd1) - Okta - Bengaluru, India
- [Staff Software Engineer, Inference / Compute Infrastructure Engineering](https://hotfix.jobs/jobs/cf3c46c0-63dd-4681-b117-16200b400d89) - Together AI - London, United Kingdom
- [Staff Software Engineer](https://hotfix.jobs/jobs/17a92e38-a4f6-4c6b-a6cc-e3bf61bd9955) - Okta - Bengaluru, India
- [Senior/Staff Kubernetes Infrastructure Engineer](https://hotfix.jobs/jobs/5276ef82-8104-4269-8e3f-7f0e02d35c2d) - Fal - Remote - $180k – $250k/yr

**Apply:** https://hotfix.jobs/jobs/59c5d952-2d02-4bf3-828f-ba895f4faf84
**Canonical:** https://hotfix.jobs/jobs/59c5d952-2d02-4bf3-828f-ba895f4faf84