Site Reliability Engineer (Senior or Staff), Infrastructure Security
Senior or Staff Site Reliability Engineer leads design and implementation of cloud security solutions (AWS, Azure, GCP), builds automation for monitoring and alerting, and mentors SRE team. Requires 6+ years SRE/infra experience with security focus, IaC proficiency, and cloud expertise.
About the job
Responsibilities
Cloud Security Design and Implementation:
- Help lead the design and deployment of security solutions for cloud platforms (AWS, Azure, GCP), including network and compute security, identity management, and cloud security posture management (CSPM)
Automation and Monitoring:
- Build automated solutions for real-time security monitoring, logging, and alerting in cloud environments. Leverage native cloud services and third-party tools for runtime security monitoring and anomaly detection
Security Tooling:
- Evaluate, implement, and manage cloud-native security tools and platforms for endpoint security, identity management (IAM), and CSPM
Qualifications
Experience:
- 6+ years of experience in SRE, infrastructure engineering or similar role, with a strong focus on security work, with ideally 2+ years in a senior or staff engineering role
Security Mindset:
- A comprehensive understanding of all facets of cloud environment security, spanning from foundational OS networking layers to cloud provider configurations. Proven experience in leading projects within security-focused areas, such as runtime scanning, security observability, CSPM, and more
Cloud Expertise:
- Strong experience with at least one cloud platform (AWS, Azure, GCP), including expertise in IAM, VPC networking, security groups, and cloud security tools (e.g., GuardDuty, Security Hub, CloudTrail)
Coding/Automation:
- Proficiency in at least one programming language (we use Golang but are language agnostic when it comes to hiring) and experience with infrastructure-as-code tools (Terraform, CloudFormation, Ansible) to automate security configurations and processes
Linux and Networking:
- Understanding of the underlying Linux and networking concepts, including low-level fundamentals, and how they work together in complex systems
Communication and Leadership Skills:
- Strong ability to explain complex security concepts to both technical and non-technical teams. Ability to lead a small technical team and ensure success both meeting the team goals as well as personal growth for all team members
Skills
AWS, Azure, GCP, Terraform, CloudFormation, Ansible, Go, IAM, Cspm, Guardduty, Linux, Networking
Similar jobs
DevOps / SRE jobsOwn and scale Nango’s cloud platform, customer-controlled deployments, infrastructure automation, reliability, and data layer. The role requires 10+ years in platform, infrastructure, DevOps, or SRE work, with deep Kubernetes, AWS, Terraform, database, and compliance experience.
Own and scale the company’s cloud platform, BYOC deployments, infrastructure automation, reliability, data layer, and infrastructure security. Requires 10+ years in platform, infrastructure, DevOps, or SRE roles, with deep Kubernetes, AWS, Terraform, and database expertise.
Leads the architecture, automation, observability, and reliability of multi-region AWS infrastructure supporting mission-critical payment systems. Requires 8+ years of distributed-systems experience and deep expertise in infrastructure as code, Kubernetes, automation, and cloud networking.
Leads the design and deployment of AI-enabled manufacturing systems, MES, connected-factory infrastructure, and automation for aircraft production. Requires a bachelor’s degree and 8+ years of experience in digital manufacturing, industrial automation, or software-enabled operations.
Designs, automates, and operates AWS infrastructure, shared development environments, and container platforms. The role requires strong experience with Kubernetes, infrastructure as code, environment lifecycle automation, cloud security, compliance, and cost optimization.