Skip to content
WatershedWatershed

Software engineer, cloud infrastructure

Builds and maintains cloud infrastructure systems on Google Cloud to support engineering teams in deploying, scaling, and observing production workloads. Requires 3+ years experience in infrastructure engineering.

About the job

The role

The Cloud Infrastructure team owns foundational capabilities and tools that engineering teams use to orchestrate workloads, deploy, test, scale, and observe code in production. We're looking for engineers experienced in managing cloud infrastructure at scale to empower teams and deliver high-quality user experiences.

You might be a good fit if:

  • You have 3+ years of engineering experience
  • You’ve worked on Infrastructure or Platform teams, or significant projects involving infrastructure components (e.g. multi-region architecture, CI/CD, infrastructure as code, kubernetes, release management, observability, cloud security)
  • You enjoy building appropriate tooling for company stage
  • You enjoy collaborating with engineering teams for customer outcomes

Location: San Francisco or New York office, 4 days per week.

Skills

GCP, Kubernetes, Infrastructure As Code, CI/CD, Observability, Cloud Security, Multi-Region Architecture, Release Management

Benchling

Benchling

San Francisco, CA
Software Engineer, Platform
$173k+/yrHybrid4+ YOEDevOps / SRE

Build developer-experience tooling and release systems within Benchling’s Platform team, helping engineering teams develop, test, package, and ship high-quality software rapidly. The role requires 4+ years of software engineering experience, web framework expertise, strong problem-solving, and effective cross-functional communication.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Ramp

Ramp

New York, NY
TLM, Production Engineering
$168k+/yrHybrid3+ YOEDevOps / SRE

Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.

Roboflow

Roboflow

New York, NY
Infrastructure Engineer
$165k+/yrRemoteDevOps / SRE

Infrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.

Baseten

Baseten

San Francisco, CA
Software Engineer - Continuous Delivery
$165k+/yrHybridDevOps / SRE

Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.