Senior Cloud Engineer - Product Metrics
Senior Cloud Engineer responsible for designing, operating, and improving petabyte-scale Product Metrics systems built with Golang, Kubernetes, and ClickHouse. The role requires 5+ years of experience with scalable distributed systems and production ownership.
About the job
Responsibilities
- Help determine the roadmap for the Product Metrics team.
- Design, build, operate, and maintain business-critical petabyte-scale systems.
- Own the performance, reliability, availability, and cost efficiency of Product Metrics systems.
- Deliver and iteratively improve new features in collaboration with the team.
- Mentor and support team members, participate in design discussions, and collaborate across engineering teams.
- Participate in the on-call rotation and take ownership of operated services.
Requirements
- 5+ years of relevant software development industry experience building and operating scalable, fault-tolerant, distributed systems.
- 2+ years of software application development experience using Golang.
- Experience with at least one major cloud service provider, such as AWS, GCP, or Azure.
- Experience storing, shipping, and retrieving large volumes of data efficiently using technologies such as ClickHouse.
- Experience with Kubernetes, Helm, ArgoCD, Temporal, and infrastructure-as-code tools such as Terraform.
- Strong production debugging and problem-solving skills.
- Excellent communication and teamwork skills, including in a fully remote environment.
Nice-to-haves
- Experience writing Kubernetes operators or controllers.
- Experience with Kafka streaming technology.
- Additional experience with ClickHouse.
Compensation and Benefits
- Equity through company stock options.
- Healthcare contributions.
- Flexible time off in the United States and generous entitlement in other countries.
- USD $500 home-office setup allowance for remote employees.
- Opportunities to attend company-wide offsites.
Skills
Go, Kubernetes, ClickHouse, AWS, GCP, Microsoft Azure, Helm, Argo CD, Temporal, Terraform, Kafka, Distributed Systems
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Senior software engineer building standardized, self-service cloud infrastructure across AWS, Google Cloud, and networking systems. Requires 5+ years of software engineering experience, production cloud infrastructure expertise, and proficiency in Go or Python.
Designs and supports physical IT infrastructure across offices, labs, manufacturing facilities, and data centers, including racks, cabling, power, cooling, documentation, and capacity planning. Requires 5+ years of physical infrastructure engineering experience and strong cross-functional project execution.