Staff Software Engineer, Edge
Staff Software Engineer leading automation traffic management (bots/crawlers) and rate limiting at Pinterest's Edge (CDN, TLS, DNS, proxies). Design/implement Envoy-based L7 logic in C++/Go/Python; own roadmap, mentor, and drive reliability for 600M+ user platform.
About the job
What you’ll do
- Build Edge systems that combine in-house implementation, open-source software like Envoy proxy, and multiple CDN & DNS providers. Balance what to build, what to leverage from the community, and what to orchestrate across vendors.
- Drive automation traffic management efforts in the Edge space, designing solutions to manage bots and crawlers at scale.
- Design and implement request processing logic in Envoy to support Edge capabilities such as routing, traffic classification, and enforcement. Write, review, and maintain code written in C++, Golang, and Python.
- Collaborate with CDN vendors and internal platform teams to define integration points.
- Study internal user pain points with the rate limiting platform and translate findings into a prioritized roadmap.
- Mentor engineers within the team and demonstrate technical leadership through design reviews, pairing, and coaching.
- Use AI to accelerate prototyping, analysis, and operational tasks, while applying engineering judgment to verify correctness and production-readiness.
- Participate in oncall for Edge systems and drive incident response, post-mortems, and reliability improvements.
What we’re looking for
- 6+ years of experience in infrastructure, networking, or platform engineering, with a deep focus on traffic management systems (CDN, HTTP gateways, service mesh, or load balancing).
- Experience with bot detection, scraping prevention, or automation traffic management systems is a plus.
- Proficiency in C++, Golang, or Python, with experience developing L7 HTTP proxies such as Envoy, Nginx, Varnish, and Traefik.
- Deep understanding of network protocols across L3-L7 and hands-on experience debugging production networking issues at scale.
- Experience gathering requirements from internal users, aligning cross-functional stakeholders, and delivering credible technical plans.
- Proven ability to design distributed systems architectures and drive them from proposal through proof of concept to production.
- Demonstrated experience using AI to accelerate engineering workflows (design exploration, code generation, operational analysis), with a clear approach to validating correctness and quality.
- Bachelor's, Master's, or PhD degree in Computer Science, Computer Engineering, or a related field, or equivalent experience.
Skills
C++, Go, Python, Envoy, Nginx, Varnish, Traefik, Cdn, DNS, Rate Limiting, L7 Http Proxies, Network Protocols L3-L7, Distributed Systems
Similar jobs
DevOps / SRE jobsBuild and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.
Leads strategic production engineering initiatives that improve the reliability, scalability, observability, and security of large-scale platforms. The role requires 7+ years of relevant experience, strong coding skills, and expertise in reliability practices such as SLIs, SLOs, and incident management.
Leads the operational reliability, security, observability, deployment standards, and governance of Databricks for enterprise data workloads. Requires 12+ years in platform, SRE, or cloud data infrastructure engineering plus production Databricks experience and expertise in CI/CD, secure execution, and regulated environments.
Build and operate secure, highly available Kubernetes platforms on AWS, including cluster creation, scaling, service mesh, automation, and incident response. The Staff-level role requires deep experience with Kubernetes, Terraform, AWS, Helm, Karpenter, and Istio.
Leads reliability and networking for highly available, secure cloud services in Okta’s Federal SRE organization. The role requires active TS/SCI clearance with full-scope polygraph, Federal/DoD compliance experience, and deep expertise in AWS networking, Terraform, observability, and automation.