Staff Infrastructure Engineer, Trading
Own the infrastructure, deployment, and operational tooling for Coinbase’s latency-sensitive institutional trading platform across cloud and colocated environments. The role requires 8+ years of infrastructure, platform, or SRE experience, strong Linux and networking fundamentals, and experience operating regulated, low-latency systems.
About the job
Responsibilities
- Own infrastructure, deployment, and operational tooling for latency-sensitive, multi-node trading environments across cloud and on-premises/colocated venues.
- Drive reliability and developer velocity through observability, deployment safety, and incident response for 24/7 systems.
- Establish operational standards, reviews, and automation while reducing single points of failure across the trading stack.
- Partner with trading-platform engineers to make latency-sensitive systems operable and performant.
- Mentor engineers and build team resilience in a lean, high-impact environment.
- Partner with Product, Institutional Markets, and SRE to turn platform needs into a roadmap.
Requirements
- 8+ years of infrastructure, platform, or SRE engineering experience with ownership of production systems at scale.
- Experience running infrastructure for latency-sensitive trading environments, including on-premises/colocated deployments.
- Proficiency with orchestration tooling, containerized and bare-metal deployments, and observability/logging tooling.
- Strong Linux performance and networking fundamentals in latency-critical environments.
- Experience delivering end-to-end infrastructure solutions, including scoping, implementation, deployment safety, monitoring, and incident response.
- Experience in a regulated or financial environment where reliability and low-latency performance are critical.
- Responsible use of generative AI with human oversight.
Compensation and Benefits
- Annual base salary: $218,025–$256,500 USD.
- Total compensation may include equity and bonus eligibility.
- Benefits include medical, dental, vision, and 401(k).
Skills
Linux, Networking, Kubernetes, Containerization, Bare-Metal Deployments, Observability, Logging, Incident Response, Deployment Automation, Cloud Infrastructure
Similar jobs
DevOps / SRE jobsLeads development of Coinbase’s CI, build, and deployment infrastructure used by engineers across the organization. The role requires 8+ years building production distributed systems, strong Go or systems-language expertise, and demonstrated technical leadership across complex platform initiatives.
Leads reliability engineering for Reddit’s critical user-facing systems, improving availability, scalability, performance, automation, and incident response at internet scale. Requires 8+ years operating distributed systems and strong expertise in programming, observability, high availability, and production troubleshooting.
Provides technical leadership for reliability, scalability, and operational excellence across Reddit’s advertising systems. The role requires 8+ years operating large-scale distributed systems, strong software engineering skills, and expertise in cloud-native architectures, observability, and incident response.
Build and operate Reddit’s internet-scale observability platform across monitoring, logging, and distributed tracing. The role requires 7+ years of infrastructure or software engineering experience, distributed systems expertise, and strong Kubernetes and troubleshooting skills.
Leads the establishment and maturation of SRE practices across cloud infrastructure and platform services, improving observability, resilience, incident response, and operational tooling. Requires 7+ years of experience, major-cloud infrastructure expertise, infrastructure as code, distributed systems, and strong technical leadership.