Skip to content

Staff Backend Engineer - Databases Tempo | US | Remote

Leads technical direction for Tempo distributed tracing backend, owning architecture of ingestion, storage, query, and metrics generation. Drives operational excellence at scale, API design for AI/agents, and mentors engineers on complex initiatives.

About the job

What You’ll Be Doing

  • Lead multi-quarter technical initiatives from problem framing through rollout, e.g., trace aggregation APIs, Limitless Tempo, autoscaling cells and customer limits, or query engine improvements.
  • Own the architecture of core Tempo components: ingestion, storage, query, and metrics generation. Drive design reviews, make sharp trade-offs on performance, cost, and complexity, and document the “why” for the team.
  • Design APIs for humans and agents. Shape the next generation of Tempo’s interfaces (structured, deterministic, discoverable) so that Act 3 products, LLM-driven assistants, and external integrators can build on Tempo reliably.
  • Drive operational excellence. Own outcomes against concrete SLOs (P99 write latency, incident recurrence, TCO per ingested GB) and push the team toward Zero Ops through automation, parameterized rollouts, and actionable alerts.
  • Partner with Product and sibling teams. Work closely with PMs and with App Observability, Asserts, Drilldown, and Grafana Assistant teams to understand how Tempo gets consumed and to ship what unblocks them.
  • Mentor engineers. Raise the engineering bar through code review, design feedback, pairing on hard problems, and writing that leaves the team smarter than you found it.
  • Participate in on-call for the services you help build, and be a force multiplier in incident response and post-incident learning.
  • Contribute to open source. Tempo is OSS. You will engage the community, review external contributions, and help steer the project in the open.

What Makes You a Great Fit

  • Technical leadership. A track record of leading complex, multi-quarter initiatives that spanned design, delivery, and operations, and made the teams around you better.
  • Deep systems experience. Substantial hands-on experience building and operating distributed data systems in production: ingestion pipelines, storage engines, query execution, or similar.
  • Strong software craftsmanship. You write clean, robust, performant software that others can maintain, and you know when to optimize vs. when to ship.
  • Strong Go, or a path to it. We write Tempo in Go. Deep experience in other systems languages (Rust, C, C++) translates well.
  • Operational mindset. You’ve owned production services, carried a pager, reduced toil, and treated SLOs as a product feature, not a chore.
  • Customer focus and pragmatism. You break complex problems into short feedback loops: analyze, design, deliver an MVP, learn, iterate.
  • Leadership through writing and collaboration. You lead through design docs, reviews, and shipped code, not hierarchy. You communicate clearly in a fully remote, asynchronous environment.

Bonus Points For

  • Experience with tracing, OpenTelemetry, or large-scale observability systems.
  • Experience designing query languages, SQL/TraceQL-like engines, or APIs intended to be consumed programmatically (by services or agents).
  • Experience with columnar storage formats (e.g., Parquet) or purpose-built on-disk formats for analytical workloads.
  • Experience operating multi-tenant, multi-cell SaaS infrastructure at scale on Kubernetes.
  • Experience building for AI/LLM consumers: structured APIs, metadata/discovery endpoints, deterministic outputs, evaluation harnesses.
  • Open-source contribution or maintainership, and comfort engaging a community in the open.
  • Experience as an on-call user of Grafana, Prometheus, Loki, or Tempo in a previous role (or on a homelab).
  • Experience in a fully remote, globally distributed...

Skills

Go, Kubernetes, OpenTelemetry, Grafana, Prometheus, Distributed Systems, Tracing, SLOs, Query Engines, Storage Engines

Pinterest

Pinterest

Palo Alto, CA
Staff Software Engineer, Ads & Core Serving Platform
$208k+/yrHybrid8+ YOEBackend Engineering

Leads architecture and implementation of large-scale, low-latency serving platforms supporting Ads and Core experiences. The role requires 8+ years of backend or distributed-systems experience, strong technical leadership, and expertise in scalable production systems.

Airbnb

Airbnb

United States

Senior Staff Engineer, Communication & Connectivity
$248k+/yrRemote12+ YOEBackend Engineering

Senior technical IC focused on backend architecture and AI technologies for Airbnb's communication and connectivity platform. Partners with senior leaders and contributes code while providing technical leadership across teams.

Axion

Axion

New York, NY
Staff Software Engineer, Data Intelligence
$230k+/yrHybrid8+ YOEBackend Engineering

Staff Software Engineer leading backend and data-intensive systems, including scalable services, AI-powered workflows, cloud infrastructure, and reliability initiatives. Requires 8+ years of software engineering experience and expertise in distributed systems, data processing, and containerized applications.

Anthropic

Anthropic

San Francisco, CA

Staff+ Software Engineer, ML Sampling Path
$320k+/yrHybrid8+ YOEBackend Engineering

Design and operate high-QPS backend systems on Claude’s token-generation path, owning latency, reliability, safe deployments, and incident response. The role requires 8+ years of software engineering experience, strong distributed-systems expertise, and production ownership of mission-critical services.

GitLab

GitLab

United States
Staff Backend Engineer, Database Automation
$153k+/yrRemote7+ YOEBackend Engineering

Build and operate Go-based automation and a PostgreSQL-as-a-service platform for high-throughput, always-on production systems. The role requires deep PostgreSQL production experience, backend development expertise, infrastructure automation, and staff-level technical leadership.