Skip to content
DISQODISQO

Senior Site Reliability Engineer

Senior Site Reliability Engineer building agentic AI platforms, MCP servers, and automation tools to achieve zero KTLO. Drives SecDevOps culture with focus on security, reliability, observability while partnering on product roadmap and incident response. Requires 6+ years SRE/DevOps experience plus deep expertise in AWS, EKS, Terraform, Kubernetes.

About the job

What you will do

  • Partner with a team of high-performing engineers and developers who are focused on delivering best in class software products
  • Function as a Change Agent to introduce and evangelize our shift to a SecDevOps culture, solving for security, reliability, cost-effectiveness, and observability
  • Building Zero trust security designs and adapting DevSecOps mindset when building new services
  • Cultivate an automation-first attitude and work to champion code-centric solutions throughout our department to improve velocity and deliverability
  • Participate in the design process, representing, solving and planning for security, reliability and cost effectiveness prism in the product roadmap
  • Collaborate and communicate with cross-functional colleagues belonging to the same job family to drive SecDevOps centric cross-org initiatives, tooling and standards
  • Participate in incident response in collaboration with application owners and the platform team
  • Build tools that empower capacity planning and demand forecasting, software performance analysis, and system tuning
  • Design and build agentic AI platforms and AI agents — including home-grown MCP (Model Context Protocol) servers — to drive our organization toward a 0 KTLO north star, shifting engineering effort away from routine operational maintenance and toward high-value, forward-looking initiatives
  • Develop tooling that enable our product teams to through self service
  • Think implementing long-term solutions that are complete mechanisms by building tools, driving adoption and inspecting results for tuning

What you bring to the role

  • At least 6+ years of experience as SRE, DevOps or equivalent engineering roles
  • Hands-on experience building AI SRE/DevOps Agents and using AI agentic frameworks
  • Hands-on experience building internal MCPs for AI Agentic use
  • Strong hands-on experience working with foundational AWS services
  • Expert level experience working with AWS EKS
  • Strong hands-on hands-on experience with ArgoCD
  • Strong hands-on experience Terraform
  • Experience with at least one language for automation - Bash, Python, Golang, etc.
  • Experience deploying Serverless application using SAM
  • Experience building and managing CI/CD pipelines
  • Strong hands-on experience in Linux architecture, microservices and container orchestration (Docker, Kubernetes, etc.)
  • Experience with monitoring tools (New Relic, Prometheus, Grafana, Loki etc.)
  • Experience working on infrastructure projects in an Agile environment
  • Great communication, collaboration and presentation skills
  • Ability to team up with people from different disciplines and drive for a win-win
  • Ability to adapt and build security orchestration and automation at scale

Skills

AWS, EKS, Argo CD, Terraform, Python, Go, Sam, CI/CD, Linux, Docker, Kubernetes, New Relic, Prometheus, Grafana, Loki

Commure

Commure

Mountain View, CA
Senior Software Engineer, Infrastructure
$170k+/yrHybrid6+ YOEDevOps / SRE

Own foundational cloud infrastructure and the internal developer platform supporting Commure’s engineering teams. The role requires 6+ years of infrastructure, platform, or SRE experience and hands-on expertise across Kubernetes, infrastructure as code, GitOps, observability, and cloud environments.

Kindred

Kindred

United States
Senior Infrastructure Engineer
$170k+/yrRemote5+ YOEDevOps / SRE

Leads cloud infrastructure, platform strategy, deployment pipelines, and infrastructure automation for a growing consumer platform. Requires 5+ years in infrastructure, DevOps, platform engineering, or SRE, plus deep AWS, coding, containerization, and infrastructure-as-code experience.

Creditgenie

Creditgenie

Plymouth Meeting, PA
Senior DevOps/SRE Engineer
$170k+/yrOn-site5+ YOEDevOps / SRE

Own reliability, deployments, observability, compliance, and AI infrastructure across AWS and Kubernetes for a fintech platform. The role requires strong DevOps/SRE depth, backend software engineering experience, and hands-on ownership of SOC 2 and PCI-DSS controls.

Reality Defender

Reality Defender

New York, NY

Senior Dev Ops Engineer
$170k+/yrOn-site7+ YOEDevOps / SRE

Own and evolve secure, highly available AWS and Azure infrastructure, including Terraform automation, Kubernetes, CI/CD, observability, networking, and incident response. The role requires 7+ years of DevOps or related experience and strong cross-functional partnership across engineering and security.

VSCO

VSCO

San Francisco, CA

Senior Software Engineer, Infrastructure
$165k+/yrHybrid5+ YOEDevOps / SRE

Own and evolve VSCO’s AWS/EKS platform, including infrastructure as code, GitOps, CI/CD, observability, networking, and production reliability. The role requires 5+ years of hands-on infrastructure or SRE experience and strong Kubernetes, Terraform, and AWS expertise.