Build and operate production-critical GitOps deployment platforms, shared service tooling, and infrastructure automation in Go and TypeScript. The role requires 8+ years of software engineering experience plus expertise with Argo CD, Helm, Kubernetes, cloud platforms, and scalable APIs.
140k – 220k/yr
Remote8+ YOEDevOps / SRE
About the role
Responsibilities
Own and operate a production GitOps deployment API and CLI built with TypeScript, Node.js, and NestJS.
Maintain and extend self-service deployment tooling using Argo CD ApplicationSets and Helm for rolling services out to GCP Cloud Run.
Manage deployment manifest schemas, Helm charts, and ApplicationSet integrations.
Own a shared internal TypeScript package ecosystem in an Nx monorepo used across dozens of services.
Maintain the application ingress and edge-routing layer and its Helm charts.
Build high-performance internal tools and APIs in Go to manage infrastructure metadata and lifecycle.
Design long-running workflows with durable execution frameworks such as Temporal, orchestrating tasks across Git, cloud providers, and CI/CD pipelines.
Develop Model Context Protocol (MCP) servers and agentic AI workflows to automate infrastructure configuration creation, upgrades, and audits.
Requirements
8+ years of software engineering experience or equivalent, including building and operating infrastructure, platform, or delivery systems used by other teams.
Experience with delivery or GitOps tooling, including Argo CD, ApplicationSets, Helm, GitHub Actions, and declarative Git-driven deployment.
Experience building backend services and shared libraries at scale.
Proficiency in Go or TypeScript/JavaScript, with experience in or willingness to learn the other.
Understanding of software architecture, design patterns, and scalable RESTful APIs.
Experience with Kubernetes, serverless runtimes such as GCP Cloud Run, and major cloud providers including GCP and AWS.
Ability to take ownership of unfamiliar, production-critical systems, stabilize them, and document them effectively.
Nice-to-haves
Experience operating a deployment platform used by other teams.
Experience with Temporal or similar workflow engines such as Cadence or Airflow.
Proficiency with Terraform and modular, reusable modules.
Practical experience with agentic AI, MCP, or A2A frameworks.
Compensation and Benefits
Base salary: $140,000–$220,000 USD.
Additional compensation, including bonus, commission, equity, and benefits, may apply.
Comprehensive benefits and holistic mind, body, and lifestyle programs are offered.
Site Reliability Engineer modernizing a multi-cloud (AWS/Azure/GCP) environment into a scalable, observable Kubernetes-based platform using DevOps/SRE practices, AIOps, IaC, and AI-driven automation to support scientific and clinical research programs. Requires 6+ years SRE/DevOps experience with strong Linux, IaC, observability, and scripting skills.
140k – 155k/yrOn-site6+ YOEDevOps / SRE
DevOps Engineer
Pump.coSan Francisco, CA
Hands-on DevOps role owning AWS infrastructure, building developer tooling, and driving technical roadmap at an early-stage YC startup. Requires 6+ years infra/DevOps experience and strong AWS/K8s/Terraform skills.
140k – 200k/yrOn-site6+ YOEDevOps / SRE
Senior Network Systems Engineer
ForterraEast Palo Alto, CA +2
Deploys, operates, and troubleshoots network infrastructure including routers, switches, Linux appliances, and AWS resources for edge-deployed communications in DDIL environments. Requires 5+ years network engineering experience, Linux proficiency, IaC automation, and 50% domestic travel.
140k – 185k/yrHybrid5+ YOEDevOps / SRE
Senior Infrastructure Engineer
ScrunchNew York, NY +16
Senior Infrastructure Engineer designs, builds, and operates cloud infrastructure, developer tooling, observability, and reliability systems at scale, primarily on GCP. Requires high-velocity dev experience, IaC, database scaling, workflow orchestration, and production Python coding.
140k – 200k/yrRemoteDevOps / SRE
Senior Infrastructure Engineer - Postgres
ClickhouseUnited States
Senior Infrastructure Engineer owns reliability, operations, and automation for ClickHouse's Postgres integration across multi-cloud environments. Requires 7+ years SRE/DevOps experience, Postgres expertise, Terraform/Kubernetes proficiency, and strong Go skills.