Senior Cloud Data Infrastructure Engineer
Build and operate scalable cloud-native infrastructure for ClickHouse’s serverless database platform, including Kubernetes-based management, metrics systems, and distributed data-plane capabilities. Requires 5+ years of software development experience and production expertise with cloud platforms and Go, C++, or Java.
About the job
Responsibilities
- Build a cloud-native database platform on public cloud infrastructure.
- Develop and maintain infrastructure for a serverless ClickHouse cloud environment.
- Work on an in-house Kubernetes operator for seamless infrastructure management.
- Improve the metrics pipeline and build systems that generate statistics and recommendations.
- Collaborate with the ClickHouse core development team and data-plane teams on infrastructure use cases and internal improvements.
- Architect and build robust, scalable, highly available distributed infrastructure.
- Participate in PagerDuty on-call rotations, debug production issues, and solve complex operational problems.
Requirements
- 5+ years of relevant software development experience building and operating scalable, fault-tolerant, distributed systems.
- Production experience with Go, C++, or Java.
- Expertise with a public cloud provider such as AWS, Google Cloud, or Azure and infrastructure-as-a-service offerings such as Amazon EC2.
- Experience with data storage, ingestion, and transformation tools such as Apache Spark or Apache Kafka.
- Experience working with distributed systems.
- Strong problem-solving, communication, and cross-functional collaboration skills.
- Willingness to participate in production on-call support.
Benefits
- Flexible work environment for remote employees.
- Employer healthcare contributions.
- Company stock options.
- Flexible or generous time off, depending on country.
- USD $500 home-office setup stipend for remote employees.
- Opportunities to attend company-wide offsites.
Skills
Kubernetes, Go, C++, Java, AWS, GCP, Azure, Amazon Ec2, Spark, Apache Kafka, Distributed Systems, Pagerduty
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Build and mature Mozilla’s internal developer infrastructure platform, including CI/CD, observability, Kubernetes optimization, environment bootstrapping, and cost optimization. The role requires 5+ years of software engineering experience, cloud-native expertise, and strong technical leadership.
Senior Software Engineer building and improving Mozilla’s internal developer infrastructure platform, including CI/CD, observability, Kubernetes, cloud optimization, and developer productivity workflows. Requires 5+ years of software engineering experience and expertise in cloud-native or platform engineering.