Senior Cloud Engineer - Product Metrics
Design, build, and operate petabyte-scale Product Metrics systems that process massive event volumes with strong reliability, performance, and availability. Requires 5+ years of distributed-systems experience, Golang expertise, cloud-platform experience, and familiarity with Kubernetes and infrastructure as code.
About the job
Responsibilities
- Help define the Product Metrics team roadmap.
- Design, build, operate, and maintain business-critical, petabyte-scale distributed systems.
- Deliver new features and iteratively improve existing systems.
- Own the performance, reliability, availability, and cost efficiency of Product Metrics systems.
- Mentor teammates, participate in design discussions, and collaborate across engineering teams.
- Participate in the on-call rotation and take ownership of operated services.
Requirements
- 5+ years of relevant software development experience building and operating scalable, fault-tolerant, distributed systems.
- 2+ years of software application development experience with Golang.
- Experience with at least one major cloud service provider, such as AWS, Google Cloud, or Azure.
- Experience storing, shipping, and retrieving large volumes of data efficiently using technologies such as ClickHouse.
- Experience with Kubernetes, Helm, Argo CD, Temporal, and infrastructure-as-code tools such as Terraform.
- Strong production debugging and problem-solving skills.
- Excellent communication and teamwork skills in a fully remote environment.
Nice-to-haves
- Experience with ClickHouse.
- Experience writing Kubernetes operators or controllers.
- Experience with Kafka streaming technology.
Compensation and Benefits
- Employer healthcare contributions.
- Company stock options.
- Flexible time off in the United States and generous time-off entitlements in other countries.
- USD $500 home-office setup allowance for remote employees.
- Opportunities to participate in company-wide offsites.
Skills
Go, Kubernetes, ClickHouse, AWS, GCP, Microsoft Azure, Helm, Argo Cd, Temporal, Terraform, Kafka, Distributed Systems
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Senior software engineer building standardized, self-service cloud infrastructure across AWS, Google Cloud, and networking systems. Requires 5+ years of software engineering experience, production cloud infrastructure expertise, and proficiency in Go or Python.
Designs and supports physical IT infrastructure across offices, labs, manufacturing facilities, and data centers, including racks, cabling, power, cooling, documentation, and capacity planning. Requires 5+ years of physical infrastructure engineering experience and strong cross-functional project execution.