Skip to content
AnthropicAnthropic

Staff+ Software Engineer, Platform Portability

Build and operate portable infrastructure that enables Claude to run reliably across multiple cloud providers and accelerator platforms. The role requires 8+ years of distributed-systems experience, multi-cloud architecture expertise, production programming, Kubernetes, and Infrastructure as Code proficiency.

About the job

Responsibilities

  • Design and build provider-agnostic abstractions for storage, messaging, and connectivity across cloud providers.
  • Develop tooling for teams to provision, configure, and operate services through a consistent multi-cloud interface.
  • Identify and eliminate provider-specific assumptions in the codebase.
  • Own end-to-end delivery of Claude on partner cloud platforms, including build, release, and validation pipelines.
  • Collaborate with cloud partners and internal infrastructure teams to develop features, debug platform issues, and influence roadmaps.
  • Review designs for cloud-portability implications and mentor engineers on multi-cloud patterns.
  • Participate in on-call rotations and improve postmortems, runbooks, and incident response.

Requirements

  • 8+ years of experience building and operating production distributed systems.
  • Experience owning large-scale infrastructure systems and leading ambiguous, cross-team technical projects from design through production.
  • Experience running production workloads across multiple cloud providers, including migration or abstraction work.
  • Understanding of cloud IAM models and networking primitives, including VPC peering, private connectivity, and service mesh.
  • Production programming experience in a general-purpose language such as Python, Go, or Rust.
  • Experience with Kubernetes and Infrastructure as Code tools such as Terraform or Pulumi.
  • Understanding of ML infrastructure and supporting interconnects.
  • Experience shipping software through hyperscaler marketplaces or managed-model platforms.
  • Background in security, compliance, data residency, regulated environments, or sovereign clouds.
  • Bachelor's degree or equivalent combination of education, training, and experience.
  • Alignment with Anthropic's mission to build safe, beneficial AI.

Compensation and Benefits

  • Annual salary range: $405,000–$485,000 USD.
  • Competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office collaboration space.
  • Hybrid policy: staff are expected to work from an office at least 25% of the time; some roles may require more.
  • Visa sponsorship may be available.

Skills

Distributed Systems, Cloud Computing, Kubernetes, Terraform, Pulumi, Python, Go, Rust, Cloud Iam, Vpc Peering, Service Mesh, ML Infrastructure, Infrastructure As Code, Cloud Marketplaces, Data Residency

Anthropic

Anthropic

San Francisco, CA
Staff+ Site Reliability Engineer, Safeguards ML Infra
$320k+/yrHybrid8+ YOEDevOps / SRE

Staff-level site reliability engineer responsible for safely deploying and operating safeguards infrastructure across model releases and cloud platforms. The role emphasizes production change management, high-stakes incident response, and automating manual launch and validation processes.

Headway

Headway

San Francisco, CA
Staff Infrastructure Engineer
$265k+/yrRemote8+ YOEDevOps / SRE

Own the cloud platform, deployment architecture, container infrastructure, networking, autoscaling, cost controls, and Python runtime health for a high-scale healthcare technology platform. The role requires 8+ years in infrastructure, platform, or SRE work, deep AWS expertise, Terraform experience, and Staff-level cross-team influence.

Polymarket

Polymarket

New York, NY

Staff Infrastructure Engineer
$250k+/yrOn-site7+ YOEDevOps / SRE

Staff Infrastructure Engineer responsible for designing and operating scalable infrastructure for growth systems, including onboarding, referrals, and user acquisition. The role requires 7+ years of production infrastructure experience, strong reliability instincts, and independent judgment in a high-autonomy environment.

Crusoe

Crusoe

San Francisco, CA
Senior Staff Deployment Automation Engineer
$250k+/yrOn-site12+ YOEDevOps / SRE

Owns deployment, CI/CD, and integration-testing automation for large-scale multi-node GPU and CPU clusters. The role requires 12+ years of experience, strong Python or Bash skills, and expertise across Linux, Kubernetes, configuration management, GPU ecosystems, and high-performance networking.

Crusoe

Crusoe

San Francisco, CA
Senior Staff Software Engineer, DC Infrastructure
$250k+/yrOn-site7+ YOEDevOps / SRE

Leads software development for diagnostics, observability, automation, and repair tooling across large-scale GPU clusters and data center infrastructure. The role requires distributed systems and cloud-platform expertise, proficiency in Go, Python, Java, or Rust, and hands-on operational problem solving.