Senior Infrastructure Engineer - Postgres
Own reliability, automation, observability, and operations for ClickHouse’s Postgres integration across multi-cloud environments. The role requires 7+ years of infrastructure or SRE experience, strong Postgres and AWS expertise, and proficiency with Terraform, Kubernetes, and Go.
About the job
Responsibilities
- Lead reliability and operations for ClickHouse’s Postgres integration, including upgrades, patching, maintenance, and scaling.
- Design and implement provisioning, deployment, and service lifecycle automation across AWS, Google Cloud, and Azure.
- Develop infrastructure as code with Terraform and modern CI/CD tooling.
- Contribute Go-based tooling and services for automation, observability, and developer experience.
- Own monitoring, alerting, metrics, and tracing across environments.
- Drive incident management and postmortem practices.
- Collaborate with platform, networking, and product teams to improve service operability.
- Mentor and enable engineers.
Requirements
- 7+ years of experience in SRE, DevOps, or infrastructure engineering running distributed, production-grade systems.
- Strong knowledge of Postgres operations, scaling, and performance tuning.
- Deep hands-on AWS experience, with exposure to Google Cloud and Azure.
- Proficiency with Terraform, Kubernetes, and container-based infrastructure.
- Strong Go development skills or willingness to own production Go code.
- Familiarity with Prometheus, Grafana, Loki, OpenTelemetry, or equivalent observability tools.
- Deep understanding of SLOs, incident response, and continuous reliability improvement.
- Hands-on, resourceful, autonomous approach to operating and shipping impactful systems.
Compensation and Benefits
- Equity through company stock options.
- Healthcare contributions.
- Flexible time off in the United States.
- $500 USD home office setup benefit for remote employees.
- Flexible work environment at a globally distributed, remote-friendly company.
Skills
Postgres, AWS, GCP, Azure, Terraform, Kubernetes, Go, CI/CD, Prometheus, Grafana, Loki, OpenTelemetry, SLOs, Incident Response, Infrastructure As Code
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Senior software engineer building standardized, self-service cloud infrastructure across AWS, Google Cloud, and networking systems. Requires 5+ years of software engineering experience, production cloud infrastructure expertise, and proficiency in Go or Python.
Designs and supports physical IT infrastructure across offices, labs, manufacturing facilities, and data centers, including racks, cabling, power, cooling, documentation, and capacity planning. Requires 5+ years of physical infrastructure engineering experience and strong cross-functional project execution.