Skip to content
DatadogDatadog

Staff Engineer - Cloud Networks

Leads the technical direction, design, and operation of large-scale multi-cloud network infrastructure, with a focus on connectivity, reliability, performance, and cost efficiency. Requires deep BGP and software-defined networking expertise plus strong software development and production operations experience.

About the job

Responsibilities

  • Design, build, and operate cloud network infrastructure across AWS, GCP, Azure, and Neoclouds in a multi-region environment.
  • Own connectivity between clouds, customers, and developers, ensuring scalable, secure, and reliable network paths.
  • Set technical direction for expanding data centers and evolving the network while maintaining stability and performance.
  • Improve cross-site and cross-region connectivity patterns to support platform needs.
  • Lead investigations into latency, packet loss, and connectivity failures, from packet capture and path analysis through cloud-provider escalations.
  • Identify and deliver network-related efficiency and cost-saving opportunities.

Requirements

  • Deep networking expertise, including BGP, route policies, path selection, and prefix advertisement.
  • Experience designing, building, and evolving large-scale software-defined networks, including interconnected environments.
  • Strong software development skills and a code-first approach to operating networks.
  • Ability to own ambiguous problems, set direction, and drive cross-team collaboration across distributed teams.
  • Practical experience operating production infrastructure and responding to incidents, including on-call ownership where applicable.

Benefits and Compensation

  • Generous and competitive benefits package.
  • New-hire stock equity (RSUs) and employee stock purchase plan.
  • Continuous career development and pathing opportunities.
  • Employee-focused onboarding, mentoring, and cross-departmental buddy programs.
  • Healthcare, dental, parental planning, mental health benefits, 401(k) plan and match, paid time off, fitness reimbursements, and discounted employee stock purchase plan.
  • Estimated yearly salary: $244,000–$305,000 USD.

Skills

AWS, GCP, Azure, BGP, Route Policies, Software-Defined Networking, Multi-Cloud Networking, Packet Analysis, Network Infrastructure, Production Infrastructure, Python, Cloud Networking

Skydio

Skydio

San Mateo, CA
Staff Site Reliability Engineer
$240k+/yrRemote8+ YOEDevOps / SRE

Owns and scales production cloud infrastructure across Kubernetes/EKS, AWS, Terraform, CI/CD, networking, and observability. The role requires 8+ years of infrastructure experience, strong Kubernetes operations expertise, and depth in reliability or scaling challenges.

Polymarket

Polymarket

New York, NY

Staff Infrastructure Engineer
$250k+/yrOn-site7+ YOEDevOps / SRE

Staff Infrastructure Engineer responsible for designing and operating scalable infrastructure for growth systems, including onboarding, referrals, and user acquisition. The role requires 7+ years of production infrastructure experience, strong reliability instincts, and independent judgment in a high-autonomy environment.

Crusoe

Crusoe

San Francisco, CA
Senior Staff Deployment Automation Engineer
$250k+/yrOn-site12+ YOEDevOps / SRE

Owns deployment, CI/CD, and integration-testing automation for large-scale multi-node GPU and CPU clusters. The role requires 12+ years of experience, strong Python or Bash skills, and expertise across Linux, Kubernetes, configuration management, GPU ecosystems, and high-performance networking.

Crusoe

Crusoe

San Francisco, CA
Senior Staff Software Engineer, DC Infrastructure
$250k+/yrOn-site7+ YOEDevOps / SRE

Leads software development for diagnostics, observability, automation, and repair tooling across large-scale GPU clusters and data center infrastructure. The role requires distributed systems and cloud-platform expertise, proficiency in Go, Python, Java, or Rust, and hands-on operational problem solving.

Headway

Headway

San Francisco, CA
Staff Infrastructure Engineer
$265k+/yrRemote8+ YOEDevOps / SRE

Own the cloud platform, deployment architecture, container infrastructure, networking, autoscaling, cost controls, and Python runtime health for a high-scale healthcare technology platform. The role requires 8+ years in infrastructure, platform, or SRE work, deep AWS expertise, Terraform experience, and Staff-level cross-team influence.