Skip to content
IllumioIllumioSunnyvale, CA

Site Reliability Engineer II

Site Reliability Engineer II responsible for designing, deploying, and maintaining multi-cloud infrastructure (Azure primary, AWS/GCP) for Illumio's SaaS products. Focus on IaC, CI/CD pipelines, monitoring, incident response, automation, and improving reliability/scalability in collaboration with engineering and security teams. Requires 2+ years SRE/DevOps experience with Azure.

141k – 162k/yr
On-site2+ YOEDevOps / SRE

About the role

Your Impact

As an SRE Engineer II, you will be responsible for managing our multi-cloud infrastructure on Azure, AWS and/or GCP. As and when required, you will be responsible for designing new services and applications in the cloud(s) and take them from development to production while working closely with Engineering, SRE/OPS, and Security teams.

On a day-to-day basis, you will work on enhancing system reliability and scalability of Illumio SaaS products, and drive continuous improvement initiatives.

  • Design, deploy, and maintain cloud infrastructure solutions on Azure, AWS, and/or GCP to support our applications and services
  • Implement infrastructure as code (IaC) principles using tools such as Terraform, ARM templates, or CloudFormation to automate provisioning and configuration management
  • Develop and maintain CI/CD pipelines for automated software delivery and deployment, leveraging tools such as Azure DevOps, AWS CodePipeline, or Jenkins
  • Monitor system performance, application health, and infrastructure metrics using cloud monitoring and logging services, and implement proactive measures to optimize performance and availability
  • Support incident response and resolution efforts, conduct root cause analysis, implement corrective actions, and document post-incident reviews
  • Collaborate with Engineering teams to design and implement scalable and reliable architectures, providing guidance on best practices for cloud-native application development
  • Implement security best practices and controls in cloud environments to protect data, applications, and infrastructure, and ensure compliance with regulatory requirements
  • Drive automation initiatives to streamline operational tasks, reduce manual effort, and improve overall efficiency in cloud operations
  • Stay current with cloud platform updates, trends, and best practices, and evaluate emerging technologies for potential adoption to drive innovation and efficiency
  • Provide support and guidance to junior team members, fostering a culture of learning, collaboration, and continuous improvement within the SRE/DevOps team

Your Toolkit

  • Bachelor's degree in Computer Science, Engineering, or related field; or equivalent work experience
  • 2+ years of experience working as an SRE, DevOps Engineer, or similar role, with hands-on experience in Azure cloud platform in a production environment setting
  • Exposure to AWS and/or GCP cloud platforms is preferred
  • Proficiency in scripting and programming languages such as PowerShell, Python, or Go for automation and infrastructure management tasks
  • Experience with CI/CD tools and methodologies, containerization technologies, and microservices architecture in cloud environments
  • Strong analytical, problem-solving, and communication skills, with the ability to collaborate effectively with cross-functional teams
  • Azure certifications such as Azure Administrator, Azure Developer, or AWS/GCP certifications are a plus

Skills

AzureAWSGCPTerraformCloudFormationarm templatesCI/CDJenkinsPythonPowerShellGoKubernetesDockerMicroservices

Similar roles

DevOps / SRE jobs
Skydio

Software Engineer - Infrastructure

SkydioSan Mateo, CA

Infrastructure engineer responsible for maintaining and scaling Kubernetes fleets, improving CI/CD, and making product-level code changes in Python or Go to support autonomous drone platform needs.

140k – 210k/yr
Hybrid2+ YOEDevOps / SRE
Harper

Forward Deployed Engineer

HarperSan Francisco, CA

Technical generalist embeds with operations teams to identify high-impact problems and rapidly builds AI agents, automations, and tools to eliminate friction. Requires 2-5 years software engineering experience with Python/TypeScript proficiency and business impact focus.

140k – 200k/yr
On-site2+ YOEDevOps / SRE
Astronomer

Software Engineer

AstronomerNew York, NY

Software Engineer building and operating Astronomer's multi-tenant cloud platform and Astro Private Cloud. Focus on Kubernetes, IaC (Terraform), cloud networking (AWS/GCP/Azure), observability, and production reliability with on-call duties. 0-4 years experience; ideal for early-career infra engineers.

144k – 235k/yr
HybridEntry levelDevOps / SRE
Crusoe

Associate Systems Software Engineer

CrusoeSan Francisco, CA

Develops Linux-based compute applications for managing virtualization stacks across AI compute servers, integrates with AI hardware like GPUs and NICs, and optimizes performance for AI/ML workloads in datacenters. Requires Linux kernel familiarity, systems programming, and hardware integration skills.

137k – 161k/yr
On-siteEntry levelDevOps / SRE
Yext

Systems Engineer

YextNew York, NY

Design, automate, and maintain reliable infrastructure across cloud and colocation environments. Build monitoring, self-service tools, and standards for distributed systems in a Linux-heavy stack.

137k – 164k/yr
On-site2+ YOEDevOps / SRE