Platform Infrastructure Engineer
Build and operate highly scalable, reliable cloud infrastructure and the platforms, tools, and automation that support Snowflake’s globally distributed services. The role requires software engineering expertise, cloud experience, and strong skills in at least one infrastructure domain.
About the job
Responsibilities
- Build a deep understanding of Snowflake’s infrastructure and services.
- Build and operate highly scalable, resilient, and performant cloud infrastructure.
- Provide technical leadership on complex projects involving infrastructure, platforms, tools, and cloud automation.
- Drive adoption of the platform to meet business goals.
- Mentor junior team members and promote high-quality code, documentation, and software development practices.
- Optimize infrastructure reliability, availability, serviceability, performance, and cost efficiency.
- Troubleshoot and resolve complex technical issues.
Requirements
- Bachelor’s or master’s degree in computer science or equivalent experience.
- 2+ years of experience on a platform or cloud team supporting mission-critical services and infrastructure in a SaaS environment.
- Strong software engineering fundamentals, coding skills, and knowledge of software development best practices.
- 1+ years of cloud computing experience with AWS, Azure, or Google Cloud.
- Fluency in one or more of Golang, Java, Python, or C.
- Expertise in at least one of container platforms, automation, networking, operating systems, site reliability, configuration management, or infrastructure as code.
- Strong attention to detail and ability to build reliable, scalable software systems.
- Effective communication and collaboration skills.
- Ability to self-manage, drive project success, and solve complex problems.
Nice-to-haves
- Experience with Pulumi or Terraform.
Compensation
- Salary and benefits information is provided on the Snowflake Careers Site for jobs located in the United States.
Skills
AWS, Microsoft Azure, GCP, Go, Java, Python, C, Container Platforms, Networking, Operating Systems, Site Reliability, Configuration Management, Pulumi, Terraform, Infrastructure As Code
Similar jobs
DevOps / SRE jobsBuild and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Platform engineer responsible for forecasting and automating compute capacity across regions, including reservations, fleet reconciliation, observability, and cost optimization. Requires 5+ years in infrastructure, SRE, platform, or capacity engineering plus production software and AWS EC2 experience.
Provides hands-on L2 technical escalation support for enterprise customers in the APAC region, troubleshooting distributed systems and APIs while leading root-cause analysis, support process improvements, and technical documentation. Requires 4+ years of support or escalation engineering experience.
Automate, manage, and optimize large-scale ClickHouse clusters handling trillions of events and 100+ PB data. Build provisioning systems with Terraform, Ansible, Kubernetes; focus on performance, scaling, and bleeding-edge features.
The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.