Cloud Infrastructure Engineer building and operating secure identity, access, and cloud platforms across AWS and Azure for Snowflake. Requires 5+ years in IAM, identity engineering or security engineering, with strong experience in Terraform, Kubernetes, modern identity platforms, and Zero Trust principles.
176k – 253k/yr
Hybrid5+ YOEDevOps / SRE
About the role
Responsibilities
Design, build, and operate secure, scalable cloud infrastructure and identity platforms across AWS and Azure.
Implement and manage IAM, IGA, authentication, authorization, SSO, MFA, identity lifecycle management, and provisioning/deprovisioning solutions using modern identity platforms and standards such as SAML, OAuth2, OIDC, and SCIM.
Develop automation, integrations, and infrastructure-as-code solutions using Terraform and programming languages such as Python, Go, or PowerShell.
Design and implement security controls for AI-powered systems, including controlled, audited, and governed agent workflows, while contributing to core security services such as service identity, secrets management, key management, authentication, and authorization.
Partner with Security and Engineering teams to deliver secure-by-design solutions, implement Zero Trust principles, and reduce operational friction.
Write high-quality, reliable code, participate in architecture and code reviews, support critical production systems, and drive operational excellence through scalability, resiliency, and automation.
Requirements
5+ years of experience in Cloud Infrastructure, Identity & Access Management (IAM), Identity Engineering, or Security Engineering.
Hands-on experience with any Identity platforms - Okta, Microsoft Entra ID (Azure AD), and modern IAM/IGA platforms.
Strong knowledge of authentication, authorization, identity lifecycle management, and federation protocols including SAML, OAuth2, OIDC, SCIM, and RBAC.
Experience designing and operating identity and access controls across AWS and Azure environments, with experience building and operating production services on AWS, Azure, or GCP.
Experience deploying and operating services on Kubernetes.
Strong automation and coding skills with Python, Go, PowerShell, Terraform, or similar technologies.
A security-first mindset with experience implementing Zero Trust, least-privilege access, and compliance frameworks such as SOC2, FedRAMP, or ITAR.
An operational mindset with experience supporting and improving production services through monitoring, troubleshooting, incident response, and automation.
Excellent collaboration and communication skills, with a proven ability to drive projects and influence technical decisions across teams.
Build and own automation, observability, and repair pipelines for one of the world's largest GPU compute fleets at hyperscale. Requires strong production engineering experience, hardware intuition at the firmware/silicon level, on-call ownership, and fluency with AI coding tools.
175k – 300k/yr
On-site5+ YOEDevOps / SRE
Software Engineer, Cloud Infrastructure
FluidstackSan Francsisco, CA +3
Build and own the observability platform, control plane APIs, and fleet state management for a hyperscale GPU infrastructure powering AI compute at 10-100s of GW scale. Requires production service ownership at scale, comfort with AI coding tools, and on-call incident response.
175k – 300k/yr
On-site5+ YOEDevOps / SRE
Production Engineer, Network
FluidstackAustin, TX
Own end-to-end network fleet health, monitoring, debugging tooling, and automated repair pipelines for massive AI datacenter infrastructure at Fluidstack. Requires systems thinking, automation-first mindset, on-call ownership, and daily use of AI coding tools like Claude/Cursor alongside Go/Python and network protocols.
175k – 300k/yr
On-site5+ YOEDevOps / SRE
Production Engineer, Compute
FluidstackSan Francisco, CA +3
Own end-to-end health, repair automation, and qualification of a hyperscale GPU/TPU compute fleet. Build metrics pipelines, firmware tooling, and self-healing repair workflows across Kubernetes and bare metal.
175k – 300k/yr
Hybrid5+ YOEDevOps / SRE
SWE - Backend Infrastructure Engineer
SesameSan Francisco, CA +2
Builds and scales core infrastructure including ML training/serving, Kubernetes clusters, and low-latency voice/audio pipelines. Requires 3+ years in infrastructure/ML systems, hands-on reliability engineering, and Kubernetes expertise.