Platform Engineer (Remote)
Builds and maintains internal platform infrastructure including Kubernetes clusters, stateful services like databases and Bitcoin/LN nodes, and observability tools. Advises dev teams on integrations; requires strong Linux, networking, cloud, and systems programming expertise.
About the job
Roles
- Manage the underlying cloud and cluster technologies that compose our internal platform
- Manage stateful backend services such as databases and Bitcoin/LN nodes
- Utilize, operate, and extend observability tools (logging/monitoring/tracing)
- Improve robustness & security of our platform using enhancements like operators and overlay networks
- Advise/assist dev teams with integrations, such as Helm charts, cloud resources, and the k8s lifecycle
Skills
- Strong foundations in Linux, networking, security, cloud, and IaC
- Managing and troubleshooting Kubernetes clusters in production
- Operating and extending observability tools (logging/metrics/tracing)
- Experience in a systems programming language (Go, Python, Rust, etc)
Preferred
- Advanced cloud technologies: GitOps, DevSecOps, service mesh
- Database administration: PostgreSQL, etcd
- Knowledge of Bitcoin and Lightning Network
- Familiarity contributing to open source projects
Skills
Kubernetes, Linux, Go, Python, Rust, Observability, Networking, Security, Iac, Postgres, GitOps, DevSecOps, Service Mesh, Bitcoin, Lightning Network
Similar jobs
DevOps / SRE jobsBuild and maintain Cloudflare’s deployment platform, enabling progressive rollouts, health-mediated releases, and automated workflows at scale. The role requires at least four years of software development experience, backend and frontend experience, and comfort with rapid delivery and on-call support.
Own large-scale ClickHouse cluster upgrades and production operations while building tooling that improves release safety and automation. The role requires 5+ years operating stateful distributed systems, cloud and Kubernetes experience, strong debugging skills, and Go development experience.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.