Skip to content
CoinbaseCoinbase

Staff Infrastructure Engineer, Trading

Own the infrastructure, deployment, and operational tooling for Coinbase’s latency-sensitive institutional trading platform across cloud and colocated environments. The role requires 8+ years of infrastructure, platform, or SRE experience, strong Linux and networking fundamentals, and experience operating regulated, low-latency systems.

About the job

Responsibilities

  • Own infrastructure, deployment, and operational tooling for latency-sensitive, multi-node trading environments across cloud and on-premises/colocated venues.
  • Drive reliability and developer velocity through observability, deployment safety, and incident response for 24/7 systems.
  • Establish operational standards, reviews, and automation while reducing single points of failure across the trading stack.
  • Partner with trading-platform engineers to make latency-sensitive systems operable and performant.
  • Mentor engineers and build team resilience in a lean, high-impact environment.
  • Partner with Product, Institutional Markets, and SRE to turn platform needs into a roadmap.

Requirements

  • 8+ years of infrastructure, platform, or SRE engineering experience with ownership of production systems at scale.
  • Experience running infrastructure for latency-sensitive trading environments, including on-premises/colocated deployments.
  • Proficiency with orchestration tooling, containerized and bare-metal deployments, and observability/logging tooling.
  • Strong Linux performance and networking fundamentals in latency-critical environments.
  • Experience delivering end-to-end infrastructure solutions, including scoping, implementation, deployment safety, monitoring, and incident response.
  • Experience in a regulated or financial environment where reliability and low-latency performance are critical.
  • Responsible use of generative AI with human oversight.

Compensation and Benefits

  • Annual base salary: $218,025–$256,500 USD.
  • Total compensation may include equity and bonus eligibility.
  • Benefits include medical, dental, vision, and 401(k).

Skills

Linux, Networking, Kubernetes, Containerization, Bare-Metal Deployments, Observability, Logging, Incident Response, Deployment Automation, Cloud Infrastructure

Coinbase

Coinbase

United States

Staff Software Engineer, Developer Infrastructure
$218k+/yrRemote8+ YOEDevOps / SRE

Leads development of Coinbase’s CI, build, and deployment infrastructure used by engineers across the organization. The role requires 8+ years building production distributed systems, strong Go or systems-language expertise, and demonstrated technical leadership across complex platform initiatives.

Reddit

Reddit

San Francisco, CA

Staff Site Reliability Engineer - Site Experience
$217k+/yrOn-site8+ YOEDevOps / SRE

Leads reliability engineering for Reddit’s critical user-facing systems, improving availability, scalability, performance, automation, and incident response at internet scale. Requires 8+ years operating distributed systems and strong expertise in programming, observability, high availability, and production troubleshooting.

Reddit

Reddit

San Francisco, CA

Staff Site Reliability Engineer, Ads
$217k+/yrRemote8+ YOEDevOps / SRE

Provides technical leadership for reliability, scalability, and operational excellence across Reddit’s advertising systems. The role requires 8+ years operating large-scale distributed systems, strong software engineering skills, and expertise in cloud-native architectures, observability, and incident response.

Reddit

Reddit

United States

Staff Software Engineer, Observability
$217k+/yrRemote7+ YOEDevOps / SRE

Build and operate Reddit’s internet-scale observability platform across monitoring, logging, and distributed tracing. The role requires 7+ years of infrastructure or software engineering experience, distributed systems expertise, and strong Kubernetes and troubleshooting skills.

Shield AI

Shield AI

San Mateo, CA
Sr. Staff Lead Site Reliability Engineer
$220k+/yrOn-site7+ YOEDevOps / SRE

Leads the establishment and maturation of SRE practices across cloud infrastructure and platform services, improving observability, resilience, incident response, and operational tooling. Requires 7+ years of experience, major-cloud infrastructure expertise, infrastructure as code, distributed systems, and strong technical leadership.