Senior Backend Engineer (Infrastructure)
Build and scale infrastructure for high-availability, low-latency tax compliance platform handling millions of global transactions. Manage Kubernetes clusters, Postgres at scale, and cloud-native services while participating in on-call and collaborating with customers.
About the job
Responsibilities
- Find solutions to Sphere's toughest scaling, performance, and latency problems
- Work closely with engineering team to define tooling to help ship faster
- Participate in on-call rotation to solve critical production events
- Work directly with customers like Eleven Labs, Replit, Windsurf, and partners like Stripe, Chargebee on latency and availability requirements
- Influence and implement next generation of Sphere's database, real-time queue, and container orchestration infrastructure
- Introduce and scale best practices with cloud-native technologies like Amazon ALB, ECS/EKS, Temporal, AWS SQS, Amazon Aurora PostgreSQL, Elasticache Redis, and S3
- Build abstractions within Terraform to simplify architecture and increase velocity and ownership
Requirements
- Experience managing k8s clusters in AWS/GCP/Azure at scale
- Extensive experience shipping high-quality architectures for mission critical systems (focus on high availability, high load, low latency)
- Experience with Postgres at scale
Nice to Have
- Experience working with large volumes of transaction data
- Strong experience in Python (core application backend and data pipeline services built with Python and Django)
- Passionate about developer experience
- Very strong attention to detail
Skills
Kubernetes, AWS, GCP, Azure, Postgres, Python, Django, Terraform, Amazon Aurora, Elasticache Redis, Temporal, Aws Sqs, Amazon Alb, ECS, EKS
Similar jobs
DevOps / SRE jobsOwn and improve the CI/CD, testing, and deployment infrastructure that enables fast, safe, observable releases at scale. The role requires strong distributed-systems expertise, hands-on Kubernetes and infrastructure-as-code experience, and a track record of measurable cross-team improvements.
Build and evolve the developer platform that enables reliable, efficient software delivery across the company. The role requires 5+ years of software engineering experience, strong programming and system-design fundamentals, and expertise in build systems, CI/CD, testing, and deployment automation.
Own and evolve a broad infrastructure platform spanning cloud, Kubernetes, deployment, reliability, security, and GPU-backed AI systems. The role requires 8+ years operating production distributed systems, strong incident and architecture experience, and practical cloud infrastructure expertise.
Senior engineer owning safety-critical software pipelines and infrastructure, from static and dynamic analysis through CI enforcement, dashboards, and reliability tooling. Requires an advanced technical degree, 7+ years working with large codebases, and expertise in Bazel, Python, backend infrastructure, and C++.
Build and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.