Senior Software Engineer - Snowpark Container Service
Senior engineer to design, build, and lead development of Snowpark Container Services, a Kubernetes-based container compute platform. Requires 7+ years building large-scale distributed systems and strong coding skills in Java, C++, or Go.
About the job
Responsibilities
- Design and develop features, understand customer requirements and meet business goals.
- Lead a team of engineers, including mentoring and guiding them, and build technical direction and strategy for large and critical parts of the product surface area.
- Manage all aspects of the Project, including Design, Coding, Reviews, Testing, Observability, Tooling and On-Call support.
- Build highly reliable and fault-tolerant software to meet the needs of the largest customers.
- Ensure operational readiness and maintainability of the service with heavy focus on reliability, availability, debuggability and performance.
- Design and build Kubernetes based OCI compliant container compute platform features and capabilities that scale and evolve with changing business and customer needs.
Requirements
- 7+ years of industry experience building features and capabilities of large scale systems or infrastructure level platforms.
- Extremely strong fundamental computer science skills and experience building distributed systems.
- Hands-on coding experience in Java, C++ or Go is highly desirable.
- Ability to work on-site in our downtown Bellevue office.
Nice-to-Haves
- Experience building products or services with Kubernetes.
- Designing and implementing products and features with multi-cloud support (AWS, Azure, GCP).
- Implementing multi-tenant systems with focus on performance, isolation and security.
Skills
Java, C++, Go, Kubernetes, Distributed Systems, AWS, Azure, GCP, Multi-Tenant Systems
Similar jobs
DevOps / SRE jobsBuild and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.
Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.
Leads infrastructure and platform strategy for a production healthcare AI platform, owning AWS, reliability, disaster recovery, compliance, CI/CD, and secure AI-agent operations. Requires deep cloud and Terraform expertise, audit-cycle experience, and prior technical leadership.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Leads design, deployment, and operation of secure distributed cloud systems for public-sector and air-gapped environments. Requires active or obtainable TS/SCI clearance with polygraph, U.S. citizenship, and 7+ years of production experience.