Senior Cloud Engineer building secure, scalable ClickHouse Cloud platform for regulated government and enterprise environments across cloud, hybrid, and on-prem (including air-gapped). Requires 6+ years building distributed systems, Kubernetes expertise, cloud platforms, and automation with Go/Python.
Salary not listed
Remote6+ YOEDevOps / SRE
About the role
Responsibilities
Design and develop a highly available, scalable, and secure ClickHouse Cloud platform for regulated and mission-critical environments.
Build innovative deployment automation across cloud, hybrid, and on-prem systems, including disconnected environments when needed.
Work closely with existing Dataplane and Core teams to ensure software parity with existing cloud infrastructure.
Solve unique scaling, reliability, and performance challenges in regulated environments.
Design and deploy ClickHouse Cloud on Kubernetes and containerized environments ensuring high availability, replication, and backup.
Develop and maintain Helm charts, operators, and Kubernetes manifests for database management.
Implement repeatable automation to build, scale, and troubleshoot infrastructure components across diverse deployment models.
Optimize ClickHouse Cloud database performance and storage architecture for on-prem, hybrid, and government cloud deployments.
Integrate secure authentication, encryption, and access control mechanisms.
Develop and maintain technical documentation for system architecture, security, and compliance audits.
Troubleshoot and resolve database performance, security, and operational issues.
Automate deployments and lifecycle management using Terraform, Ansible, or CI/CD pipelines.
Requirements
6+ years of relevant software development industry experience building and operating scalable, fault-tolerant, distributed systems.
Experience with ClickHouse or relational (PostgreSQL, MySQL) and NoSQL (MongoDB, Cassandra) databases.
Proficiency with Kubernetes tools (Helm, Kustomize, operators, Istio, service mesh).
Experience with containerized deployments (Docker, Kubernetes, OpenShift), ideally in regulated or enterprise environments.
Experience with cloud platforms (AWS, Azure, GCP, AWS GovCloud, Azure Government, or on-prem equivalents).
Proficiency in programming/scripting languages (Go or Python) for automation and integration.
Excellent communication skills and the ability to work well within a team and across engineering teams.
Strong problem solver with solid production debugging skills.
Passionate about efficiency, availability, scalability, and data governance.
Thrive in a fast-paced environment and see yourself as a partner with the business with the shared goal of moving the business forward.
High level of responsibility, ownership, and accountability.
Nice-to-Haves
Experience with secure, regulated, or restricted network environments (including airgapped architectures).
Leads a storage engineering team while architecting and operating highly available Linux-based storage, datacenter, and data-protection infrastructure. The role requires deep Ceph experience, PB-scale archiving and backup expertise, hands-on troubleshooting, and team leadership.
215k – 245k/yrRemote5+ YOEDevOps / SRE
Senior Platform Engineer
BestowUnited States
Own platform initiatives that improve cloud scalability, reliability, automation, and developer productivity. The role requires 5+ years of cloud infrastructure experience plus hands-on expertise with infrastructure as code, Kubernetes, CI/CD, programming or scripting, and AI-assisted engineering.
145k – 171k/yrRemote5+ YOEDevOps / SRE
Senior Software Engineer, Enterprise Platform
DiscordUnited States
Build and operate Discord’s greenfield Enterprise Platform, turning identity, device management, infrastructure, and application delivery into reusable self-service software. The role requires production software engineering experience, IAM and Terraform expertise, and end-to-end ownership of complex platform projects.
196k – 221k/yrOn-site5+ YOEDevOps / SRE
Senior Site Reliability Engineer
PrizePicksUnited States
Senior Site Reliability Engineer responsible for designing, operating, and improving reliable, scalable production systems. The role requires 5+ years of reliability-focused engineering experience plus expertise in cloud platforms, infrastructure as code, Kubernetes, programming, observability, and critical incident response.
The Senior DevOps Engineer will build and operate reliable cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability across AWS environments. The role requires 5+ years in DevOps, SRE, or infrastructure engineering, with strong Terraform, Kubernetes, AWS, networking, and distributed-systems experience.