Skip to content
CodeRabbitCodeRabbitSan Francisco, CA

Senior/Staff Platform Engineer

Founding Platform Engineer owning compute, orchestration, and infrastructure for CodeRabbit's GenAI code review platform on multi-region GCP. Build from 0-to-1 with deep Kubernetes, distributed systems, and cloud infrastructure expertise.

220k – 280k/yr
Hybrid7+ YOEDevOps / SRE

About the role

Required Qualifications

  • 7+ years in Platform Engineering, Infrastructure Engineering, or Site Reliability Engineering with a strong bias toward building platforms, not just operating them
  • Deep, hands-on experience with Kubernetes: you've gone beyond running workloads to understanding and configuring the control plane, writing operators or controllers, tuning schedulers, and debugging at the runtime level; bonus if you've contributed to Kubernetes or built on its internals
  • Experience building and running large-scale distributed systems. You understand the trade-offs between consistency and availability, have debugged distributed failures in production, and have designed systems around them
  • Strong cloud compute background on GCP or AWS. You've built infrastructure, not just consumed managed services; you understand how compute, networking, and storage primitives work at the layer below the console
  • Proficiency in Docker and container runtime internals: image layering, networking modes, security contexts, and build optimization

Technical Skills

  • Container & Orchestration: CRDs, operators, admission webhooks, RBAC, network policies, autoscaling; strong Docker/OCI toolchain knowledge
  • Distributed Systems: queuing, eventual consistency, backpressure, graceful degradation
  • Infrastructure as Code: Advanced Terraform — module design, state management, programmatic provisioning patterns
  • Cloud Platforms: GCP (GKE, Cloud Run, VPC, IAM, Cloud SQL, Cloud Storage, Load Balancing) — how these work under the hood, not just how to configure them
  • Programming: Node.js/TypeScript or Go for platform tooling, operators, and automation
  • Observability: Datadog, Prometheus/Grafana or equivalent — custom instrumentation, distributed tracing, SLO-based alerting
  • Systems: Linux internals, networking fundamentals (TCP/IP, DNS, load balancing, eBPF), storage systems

Responsibilities

  • Own the compute, orchestration, and infrastructure layer that CodeRabbit's AI engine and every product on top of it run on, across a multi-region GCP footprint
  • Build the platform function from the ground up (0-to-1 role)
  • Set the technical direction, establish the patterns everyone else builds on, and make the early architectural calls
  • Build systems that need to be right, not just fast — everything else depends on them
  • High-ownership engineering culture: find problems before they're assigned, use AI as a core part of how you build, ship with judgment, and own outcomes from proposal to production

Skills

KubernetesTerraformGCPDockerGoTypeScriptNode.jsPrometheusGrafanaDatadogLinuxebpf

Similar roles

DevOps / SRE jobs
Perplexity

Member of Technical Staff

PerplexitySan Francisco, CA +2

This role is for a Software Engineer on the Cloud Infrastructure team, focusing on designing, building, and operating foundational cloud primitives and deployment models. The engineer will own the roadmap and technical strategy for agent-driven cloud infrastructure management, ensuring secure and scalable solutions for various customer environments.

220k – 405k/yr
On-site7+ YOEDevOps / SRE
Replit

Staff Infrastructure Engineer

ReplitFoster City, CA

As a Staff Infrastructure Engineer, you will ensure the reliability, scalability, and performance of Replit's infrastructure. You will drive automation, optimize performance, elevate developer experience, and mentor the engineering team on best practices for resilient systems.

220k – 325k/yr
Hybrid8+ YOEDevOps / SRE
Crusoe

Staff Software Engineer, Managed Orchestration (Managed Kubernetes)

CrusoeSan Francisco, CA +1

Staff Software Engineer designs, builds, and scales managed Kubernetes and AI training clusters, focusing on reliability, performance, and orchestration using Go, Terraform, and GCP. Oversees architecture, CI/CD pipelines, and critical infrastructure projects requiring 8+ years experience.

220k – 250k/yr
On-site8+ YOEDevOps / SRE
TruckSmarter

Staff Platform Engineer

TruckSmarterSan Francisco, CA

Staff Platform Engineer builds and owns core infrastructure platform using AWS services and IaC tools, sets architectural direction, leads security/reliability, and mentors engineers. Requires 7+ years experience with AWS, TypeScript/Node.js, and startup velocity.

220k – 280k/yr
On-site7+ YOEDevOps / SRE
Suno

Staff / Senior Software Engineer, Infrastructure

SunoCambridge, MA

Builds and operates scalable infrastructure systems including Kubernetes clusters, distributed databases, and cloud services to support AI music platform at consumer scale. Requires 5+ years experience in infrastructure engineering with strong ownership and scaling expertise.

220k – 280k/yr
On-site5+ YOEDevOps / SRE