Senior Cloud Data Infrastructure Engineer
Build and operate ClickHouse’s cloud-native database infrastructure, including Kubernetes-based management, metrics systems, and highly available distributed services. The role requires 5+ years of software development experience, production expertise in Go, C++, or Java, and experience with public cloud and data infrastructure.
About the job
Responsibilities
- Build a cloud-native database platform on public cloud infrastructure.
- Develop and maintain an in-house Kubernetes operator for seamless infrastructure management.
- Improve the metrics pipeline and build systems that generate statistics and recommendations.
- Collaborate with the ClickHouse core development team and data plane teams on infrastructure use cases and internal infrastructure improvements.
- Architect and build robust, scalable, highly available distributed infrastructure.
- Participate in PagerDuty on-call, production debugging, and incident resolution.
Requirements
- 5+ years of relevant software development experience building and operating scalable, fault-tolerant, distributed systems.
- Production experience with Go, C++, or Java.
- Experience with a public cloud provider such as AWS, Google Cloud, or Azure and infrastructure-as-a-service offerings such as EC2.
- Experience with data storage, ingestion, and transformation tools such as Spark or Kafka.
- Experience working with distributed systems.
- Strong problem-solving and communication skills, with the ability to collaborate across engineering teams.
Benefits
- Flexible work environment at a remote-friendly, globally distributed company.
- Healthcare contributions.
- Company stock options.
- Flexible time off, with country-specific entitlements.
- USD $500 home-office setup allowance for remote employees.
- Opportunities to attend company-wide offsites.
Skills
Kubernetes, Go, C++, Java, AWS, GCP, Azure, Amazon Ec2, Spark, Apache Kafka, Distributed Systems, Pagerduty
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Build and mature Mozilla’s internal developer infrastructure platform, including CI/CD, observability, Kubernetes optimization, environment bootstrapping, and cost optimization. The role requires 5+ years of software engineering experience, cloud-native expertise, and strong technical leadership.
Senior Software Engineer building and improving Mozilla’s internal developer infrastructure platform, including CI/CD, observability, Kubernetes, cloud optimization, and developer productivity workflows. Requires 5+ years of software engineering experience and expertise in cloud-native or platform engineering.