Skip to content

Platform Engineer (Remote)

Builds and maintains internal platform infrastructure including Kubernetes clusters, stateful services like databases and Bitcoin/LN nodes, and observability tools. Advises dev teams on integrations; requires strong Linux, networking, cloud, and systems programming expertise.

About the job

Roles

  • Manage the underlying cloud and cluster technologies that compose our internal platform
  • Manage stateful backend services such as databases and Bitcoin/LN nodes
  • Utilize, operate, and extend observability tools (logging/monitoring/tracing)
  • Improve robustness & security of our platform using enhancements like operators and overlay networks
  • Advise/assist dev teams with integrations, such as Helm charts, cloud resources, and the k8s lifecycle

Skills

  • Strong foundations in Linux, networking, security, cloud, and IaC
  • Managing and troubleshooting Kubernetes clusters in production
  • Operating and extending observability tools (logging/metrics/tracing)
  • Experience in a systems programming language (Go, Python, Rust, etc)

Preferred

  • Advanced cloud technologies: GitOps, DevSecOps, service mesh
  • Database administration: PostgreSQL, etcd
  • Knowledge of Bitcoin and Lightning Network
  • Familiarity contributing to open source projects

Skills

Kubernetes, Linux, Go, Python, Rust, Observability, Networking, Security, Iac, Postgres, GitOps, DevSecOps, Service Mesh, Bitcoin, Lightning Network

Cloudflare

Cloudflare

London, United Kingdom

Software Engineer: Resiliency - Deploy at Scale
No salary listedHybrid4+ YOEDevOps / SRE

Build and maintain Cloudflare’s deployment platform, enabling progressive rollouts, health-mediated releases, and automated workflows at scale. The role requires at least four years of software development experience, backend and frontend experience, and comfort with rapid delivery and on-call support.

Clickhouse

Clickhouse

Singapore
Release Engineer - Data Plane Internal Tooling and Productivity
No salary listedRemote5+ YOEDevOps / SRE

Own large-scale ClickHouse cluster upgrades and production operations while building tooling that improves release safety and automation. The role requires 5+ years operating stateful distributed systems, cloud and Kubernetes experience, strong debugging skills, and Go development experience.

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.