Staff Backend Engineer, Ads Platform
Designs and leads high-throughput, low-latency distributed backend systems powering Fetch’s advertising platform. The role requires deep distributed-systems expertise, production-scale performance and reliability experience, and technical leadership across multiple teams.
About the job
Responsibilities
- Own architecture across critical Ads services, including ad selection, targeting, pacing, delivery, attribution, and measurement.
- Design high-throughput, low-latency distributed systems with appropriate partitioning, caching, consistency, concurrency, and performance strategies.
- Build for reliability, resilience, and operational excellence under partial failures, traffic spikes, and downstream degradation.
- Establish metrics, tracing, SLOs, alerting, structured logging, capacity planning, and incident-response practices.
- Evolve event-driven and data-intensive architectures spanning transactional and non-transactional stores, streaming, messaging, caching, and real-time processing.
- Set backend, reliability, and performance standards adopted by multiple teams.
- Lead refactoring, migration, and system-health initiatives and build reusable infrastructure and patterns.
- Define and deliver multi-quarter technical initiatives across Ads, Platform, Data, and ML.
- Resolve ambiguous reliability, performance, and scalability problems spanning multiple teams.
- Mentor senior engineers, lead technical design reviews, and raise engineering standards.
Requirements
- Proven experience designing, building, and operating distributed backend systems at significant production scale.
- Deep understanding of partitioning, replication, consistency, caching, concurrency, asynchronous processing, failure isolation, and coordination.
- Strong experience diagnosing performance and scaling issues across applications, databases, caches, queues, networks, and infrastructure.
- Expertise in API design, service boundaries, event-driven architectures, streaming, messaging, distributed stores, and caching.
- Experience owning architecture for complex distributed systems, including scalability planning, performance optimization, failure modes, and long-term evolution.
- Experience building highly reliable production systems with comprehensive observability.
- Ability to evaluate tradeoffs involving scalability, reliability, latency, consistency, operational cost, complexity, and developer productivity.
- Ability to translate ambiguous product and technical problems into clear, executable direction.
Compensation & Benefits
- Full-time role available from US offices or remotely in the United States.
Skills
Distributed Systems, Backend Architecture, Event-Driven Architecture, API Design, Streaming, Messaging, Caching, Replication, Concurrency, Observability, Performance Optimization, Incident Response
Similar jobs
Backend Engineering jobsLeads architecture and implementation of large-scale, low-latency serving platforms supporting Ads and Core experiences. The role requires 8+ years of backend or distributed-systems experience, strong technical leadership, and expertise in scalable production systems.
Senior technical IC focused on backend architecture and AI technologies for Airbnb's communication and connectivity platform. Partners with senior leaders and contributes code while providing technical leadership across teams.
Staff Software Engineer leading backend and data-intensive systems, including scalable services, AI-powered workflows, cloud infrastructure, and reliability initiatives. Requires 8+ years of software engineering experience and expertise in distributed systems, data processing, and containerized applications.
Design and operate high-QPS backend systems on Claude’s token-generation path, owning latency, reliability, safe deployments, and incident response. The role requires 8+ years of software engineering experience, strong distributed-systems expertise, and production ownership of mission-critical services.
Build and operate Go-based automation and a PostgreSQL-as-a-service platform for high-throughput, always-on production systems. The role requires deep PostgreSQL production experience, backend development expertise, infrastructure automation, and staff-level technical leadership.