Skip to content
296 jobs

Job results

Figma

Figma

San Francisco, CA
Software Engineer, Traffic
$153k+/yrRemote4+ YOEDevOps / SRE

Design, build, and operate scalable distributed systems and edge networks on AWS to handle Figma's growing customer traffic and services. Requires 4+ years building infrastructure at scale, experience with TypeScript or Go, and distributed/traffic systems.

Clickhouse

Clickhouse

United States

Senior Cloud Engineer - Product Metrics
No salary listedRemote5+ YOEDevOps / SRE

Design, build, and operate petabyte-scale Product Metrics systems that process massive event volumes with strong reliability, performance, and availability. Requires 5+ years of distributed-systems experience, Golang expertise, cloud-platform experience, and familiarity with Kubernetes and infrastructure as code.

Clickhouse

Clickhouse

United States

Senior Cloud Engineer - Product Metrics
No salary listedRemote5+ YOEDevOps / SRE

Senior Cloud Engineer responsible for designing, operating, and improving petabyte-scale Product Metrics systems built with Go, Kubernetes, and ClickHouse. The role requires 5+ years of experience with scalable distributed systems and strong reliability, performance, and production debugging skills.

Supabase

Supabase

Remote

Postgres Deployment Engineer
No salary listedRemote3+ YOEDevOps / SRE

Own stability and deployment of PostgreSQL products. Package software with Nix, manage upgrades, optimize CI/CD, and resolve production issues. Requires 3+ years PostgreSQL experience and Nix proficiency.

Render

Render

United States
Software Engineer, Dev Velocity
$170k+/yrRemote5+ YOEDevOps / SRE

Build internal developer platform, tooling, and automation to accelerate engineering velocity. Focus on CI/CD pipelines, test infrastructure, build systems, and metrics to help engineers ship faster and more reliably.

Cohere

Cohere

London, United Kingdom

Senior/Staff Software Engineer, Developer Experience
No salary listedRemote5+ YOEDevOps / SRE

Builds automation and testing infrastructure for a platform, enabling reliable validation across environments and configurations. The role requires 5+ years of software engineering experience, strong Python and TypeScript skills, and expertise in CI/CD, containers, cloud platforms, and developer tooling.

Supabase

Supabase

Remote

Site Reliability Engineer
No salary listedRemote7+ YOEDevOps / SRE

SRE embedded in Service Operations to establish reliability practices, frameworks, and feedback loops across engineering teams. Focus on SLOs/SLIs, ORR processes, incident-to-improvement pipelines, and influencing without authority in a distributed environment.

Pinterest

Pinterest

San Francisco, CA

Site Reliability Engineer II
$114k+/yrRemote4+ YOEDevOps / SRE

Operate and scale a cloud-native CTV advertising platform on AWS and Kubernetes. Focus on reliability, GitOps workflows, infrastructure automation, observability, and incident response.

Capacity

Capacity

United States

Software Engineer, Enablement
$150k+/yrRemote3+ YOEDevOps / SRE

Design, build, and operate AI-powered engineering tools and developer productivity platforms. Focus on AI pairing pipelines, automated workflows, and internal tooling to accelerate engineering velocity.

Komodo Health

Komodo Health

United States

Staff Platform Engineer (Pacific Time Zone)
$196k+/yrRemote7+ YOEDevOps / SRE

Lead technical direction for Komodo's core control plane (KMC/PSS, identity, subscriptions) and App Builder/Connector. Architect platform primitives, APIs, and AI tooling in a multi-tenant SaaS environment.

Komodo Health

Komodo Health

United States

Senior Data Engineer, Sentinel (Pacific Time Zone)
$153k+/yrRemote5+ YOEDevOps / SRE

Senior Infrastructure Engineer building and operating AWS cloud infrastructure for healthcare data platform. Requires Python, Terraform, CI/CD expertise, and big data tools experience.

Chime

Chime

United States

Software Engineer, Infrastructure
$133k+/yrRemote2+ YOEDevOps / SRE

Build and operate foundational data infrastructure including Airflow, Flink, DynamoDB, and RDS using Terraform and Kubernetes. Requires 2-4 years of infrastructure/platform experience and strong Python skills.

Hightouch

Hightouch

United States

Staff Engineer, AI Productivity
$180k+/yrRemote7+ YOEDevOps / SRE

Staff-level engineer building infrastructure, tooling, and documentation to make AI coding agents dramatically more productive across the codebase. Owns agentic dev environments, MCP integrations, and agent context.

EngFlow

EngFlow

Amsterdam, Netherlands

Release Engineer
$60k+/yrRemoteDevOps / SRE

Owns weekly releases and phased production rollouts for a globally distributed remote execution platform. The role focuses on Terraform-managed infrastructure, CI health, incident coordination, operational reliability, and continuous improvement of release processes.

Scribe

Scribe

San Francisco, CA

Senior Database Reliability Engineer
$145k+/yrRemote5+ YOEDevOps / SRE

Senior IC role owning reliability, performance, and scalability of PostgreSQL (Aurora), OpenSearch, Redis, and CDC pipeline to Snowflake. Sets standards for ORM usage, migration safety, and observability at scale.

Coinbase

Coinbase

United States

Senior Site Reliability Engineer, Core AI Infrastructure
$186k+/yrRemote5+ YOEDevOps / SRE

Senior SRE owning reliability, monitoring, and automation for Coinbase's AI infrastructure on AWS and Kubernetes. Requires 5+ years cloud automation experience and strong incident response skills.

Coinbase

Coinbase

United States

Staff Site Reliability Engineer
$218k+/yrRemote8+ YOEDevOps / SRE

Staff SRE on the IT Operations team owning reliability, automation, and observability for Coinbase's AI infrastructure on AWS and Kubernetes. Requires 8+ years of cloud infrastructure experience and strong incident response leadership.

Docker

Docker

United States
Staff Software Engineer, Infrastructure
CA$238k+/yrRemote8+ YOEDevOps / SRE

Leads the design and production adoption of self-service infrastructure platforms, multi-region cloud foundations, and reliable delivery workflows. The role requires 8+ years of hands-on software or platform engineering experience, strong Go or comparable programming skills, and deep expertise in a core infrastructure domain.

Onebrief

Onebrief

United States

Staff Infrastructure Engineer
$180k+/yrRemote5+ YOEDevOps / SRE

Staff Infrastructure Engineer building and operating secure cloud-native and edge platforms for military collaboration software. Requires 5+ years production infrastructure experience, deep Kubernetes expertise, and ability to obtain SECRET clearance.

Onebrief

Onebrief

United States

Principal Infrastructure Engineer
$235k+/yrRemote8+ YOEDevOps / SRE

Principal Infrastructure Engineer building and operating secure cloud-native and edge platforms for military collaboration software. Requires 8+ years production infrastructure experience, deep Kubernetes expertise, and ability to obtain SECRET clearance.

Honor

Honor

United States

Staff Platform Engineer - Infra + DevOps
$194k+/yrRemote6+ YOEDevOps / SRE

Seasoned Platform Engineer designing and maintaining scalable distributed systems and infrastructure on AWS. Builds foundational patterns, IaC, CI/CD pipelines, and observability for Python services using Kubernetes and serverless. Requires 6+ years of platform engineering experience.

Clickhouse

Clickhouse

United States

Senior Cloud Engineer
No salary listedRemote6+ YOEDevOps / SRE

Design, deploy, and secure highly available ClickHouse Cloud platforms across regulated cloud, hybrid, on-premises, and disconnected environments. Requires 6+ years of distributed-systems experience plus expertise in Kubernetes, infrastructure automation, databases, cloud platforms, and Go or Python.

Hightouch

Hightouch

United States

Developer Productivity Engineer
$180k+/yrRemote5+ YOEDevOps / SRE

As a Senior Developer Productivity Engineer, you will own the build, test, and deployment processes for a 50+ person engineering team. You will improve monorepo productivity, drive excellence in testing, and support multi-cloud/multi-region infrastructure to enable fast and safe shipping.

LiveKit

LiveKit

NAMER
Distributed Systems Engineer
$120k+/yrRemote5+ YOEDevOps / SRE

As a Senior/Staff Distributed Systems Engineer, you will design and evolve core control, data, and observability systems for LiveKit's platform, focusing on latency, availability, and operational simplicity. You will implement resilient architectures and build tools to enhance reliability and developer velocity.

VGS

VGS

United States
Sr. Infrastructure Engineer
$145k+/yrRemote5+ YOEDevOps / SRE

As a Senior Infrastructure Engineer, you will be responsible for architecting and maintaining scalable, reliable cloud infrastructure, leading incident management, and improving operational processes. This role requires strong proficiency in AWS, infrastructure-as-code, and experience with monitoring and observability tools.

Axion

Axion

New York, NY

Senior Infrastructure Engineer
$190k+/yrRemote5+ YOEDevOps / SRE

As a Senior Infrastructure Engineer, you will design, implement, and maintain cloud infrastructure on GCP, focusing on CI/CD pipelines, Kubernetes, and Terraform. This role requires a strong background in DevOps/SRE and a passion for building foundational systems in a fast-paced environment.

Stellar Cyber

Stellar Cyber

United States

Senior DevOps Engineer/Site Reliability Engineer
$165k+/yrRemote5+ YOEDevOps / SRE

Seeking a Senior DevOps/Site Reliability Engineer to build, operate, and scale reliable cloud-native infrastructure and distributed data platforms. This role requires expertise in Kubernetes, cloud infrastructure, observability, automation, CI/CD, and incident management.

Reltio

Reltio

Lisbon, Portugal

Staff Engineer, Release Management
No salary listedRemote8+ YOEDevOps / SRE

Leads release management and cloud infrastructure initiatives for a highly available SaaS platform, mentoring the DevOps/Release team and improving automation, deployment, observability, security, and reliability. Requires 8+ years of enterprise SaaS development or operations experience and 6+ years with highly available cloud applications.

TetraScience

TetraScience

Boston, MA

Technology Lead, DevOps Engineering
No salary listedRemote7+ YOEDevOps / SRE

Hands-on technical lead owning cloud infrastructure, CI/CD pipelines, and deployment automation for a multi-tenant SaaS platform. Architect and build production systems using AWS, Terraform, CloudFormation, and Python in a GxP-regulated environment.

Chess.com

Chess.com

United States

Site Reliability Engineer
No salary listedRemote5+ YOEDevOps / SRE

Design and operate multi-regional infrastructure for a high-traffic global gaming platform, owning on-call, monitoring, automation, and hybrid cloud migration to ensure reliability at massive scale.

Order.co

Order.co

United States

Senior Site Reliability Engineer
No salary listedRemoteDevOps / SRE

Senior SRE responsible for building and operating reliable, scalable infrastructure on AWS with Kubernetes and Terraform. Focus on observability, incident response, automation, and mentoring engineers on SRE best practices.

Pinterest

Pinterest

San Francisco, CA

Staff Software Engineer, Observability
$177k+/yrRemote7+ YOEDevOps / SRE

Staff Software Engineer building and scaling Pinterest's observability platform (metrics, logs, traces) for massive distributed systems. Requires 7+ years distributed systems experience, strong data engineering skills, and expertise with modern observability tools.

Render

Render

United States
Software Engineer, Network Infrastructure
$204k+/yrRemote6+ YOEDevOps / SRE

Design, build, and operate Render's core networking stack across data centers and clouds, focusing on Kubernetes and Linux internals, traffic routing, and hybrid connectivity at scale.

ModernFi

ModernFi

New York, NY

Senior Software Engineer - Platform & Infrastructure
$160k+/yrRemote5+ YOEDevOps / SRE

Founding Senior Platform Engineer building and owning AWS cloud infrastructure, reliability, observability, security/compliance (SOC 2, Vanta), and release tooling for a fintech platform serving banks and credit unions.

Fieldguide

Fieldguide

San Francisco, CA

Staff Software Engineer, App Platform
$210k+/yrRemote10+ YOEDevOps / SRE

Lead design and evolution of core platform services, APIs, and shared primitives that power every product surface and AI agent. Drive technical standards and architecture across SaaS, enterprise, and government environments while mentoring engineers.

ClickUp

ClickUp

Poland
Senior Database Reliability Engineer
No salary listedRemote7+ YOEDevOps / SRE

The Senior Database Reliability Engineer will operate and improve large-scale PostgreSQL environments in AWS, focusing on availability, performance, security, backups, recovery, and incident response. The role requires 7+ years of database administration or engineering experience, strong Linux and SQL skills, and production cloud database expertise.

Supabase

Supabase

Remote

Software Engineer: IaC Platform Experience
No salary listedRemote5+ YOEDevOps / SRE

Owns and improves the Go-based Terraform provider for Supabase's developer platform, focusing on reliability, lifecycle management, schema evolution, and user migrations. Requires 5+ years experience with Go, deep Terraform expertise, and strong testing/CI/CD skills.

Supabase

Supabase

Remote

Edge Functions Engineer
No salary listedRemote5+ YOEDevOps / SRE

Develops and optimizes Supabase Edge Runtime, a Rust-based Deno host for global edge TypeScript functions. Evolves infrastructure for low-latency compute, integrates with Supabase stack, and improves developer tools. Requires 5+ years backend/systems experience with Rust, TypeScript, and scalable infra.

Alpaca

Alpaca

United States

Senior AI Platform Engineer
No salary listedRemote8+ YOEDevOps / SRE

Builds and maintains AI platform infrastructure for agentic systems, including connectors, execution environments, governance, and self-service tools to enable safe, scalable AI use across engineering and business teams. Requires 8+ years experience with LLM agents, GCP, and cloud-native tech.

Stellar Cyber

Stellar Cyber

Spain
Staff SRE Engineer
No salary listedRemote7+ YOEDevOps / SRE

The Staff SRE Engineer will drive reliability, scalability, observability, and operational efficiency across highly available cloud and distributed systems. The role requires 5+ years of SRE, DevOps, or platform experience, advanced Kubernetes expertise, strong automation skills, and leadership in incident management.

Railway

Railway

Remote

Senior Infra Engineer: Baremetal Orchestration
No salary listedRemoteDevOps / SRE

Builds and maintains bare metal provisioning, orchestration engine, and internal tools for Railway's infrastructure platform. Optimizes fleet efficiency and develops resilient services using Golang/Rust, Ansible, and Terraform for distributed systems.

Cerebras Systems

Cerebras Systems

Sunnyvale, CA

Member of Technical Staff (Software Engineer)
$170k+/yrRemoteDevOps / SRE

Develops and optimizes Kubernetes-based infrastructure for high-performance AI inference services, including deployment, scaling, debugging, and integration with ML workflows. Requires Master's in CS and 1+ year experience with Docker, Kubernetes, Python, and related tools.

Zoo

Zoo

Los Angeles, CA

Senior Systems Software Engineer
$145k+/yrRemoteDevOps / SRE

Builds and maintains core backend systems, infrastructure, and automation for platform reliability and scalability. Owns troubleshooting across stack, integrations with external services, and observability. Requires strong Rust, systems engineering, and distributed systems experience.

Kraken

Kraken

United Kingdom
Senior AI Compute Infrastructure Engineer
No salary listedRemote5+ YOEDevOps / SRE

This senior engineer will operate and optimize GPU and accelerator infrastructure for AI training, inference, evaluation, and experimentation. The role requires 5+ years of infrastructure experience, production GPU cluster operations, strong systems fundamentals, and expertise in serving, observability, reliability, and compute-cost optimization.

Vesta

Vesta

United States

Software Engineer - Infrastructure
$200k+/yrRemoteDevOps / SRE

Builds and scales reliable cloud infrastructure, deployment systems, observability, and developer tooling to support mortgage market operations. Requires experience with strongly typed languages, PostgreSQL, Kubernetes, and major cloud providers.

Turquoise Health

Turquoise Health

Remote

Platform Operations Engineer
$153k+/yrRemote3+ YOEDevOps / SRE

Builds and scales platform infrastructure on AWS EKS with GitOps via ArgoCD, manages CI/CD with GitHub Actions, drives observability using Datadog/Sentry/CloudWatch, and ensures reliability through SLOs and incident response. Requires 3+ years SRE/DevOps experience and Kubernetes expertise.

Tigerdata

Tigerdata

Spain
Senior Platform Engineer
No salary listedRemote3+ YOEDevOps / SRE

Builds and maintains Kubernetes-based infrastructure for managed TimescaleDB cloud services, develops Go microservices and operators, automates database operations, and ensures platform scalability and reliability. Requires 3+ years experience with Go, Kubernetes, and PostgreSQL.

Stellar Cyber

Stellar Cyber

Spain
Senior SRE Engineer
No salary listedRemote5+ YOEDevOps / SRE

Senior SRE responsible for operating and improving highly available cloud platforms, distributed data systems, observability, incident response, and deployment automation. Requires 5+ years of SRE, DevOps, or platform engineering experience, advanced Kubernetes expertise, cloud proficiency, and strong Python and Bash skills.

Railway

Railway

Remote

Senior Infra Engineer: Observability
No salary listedRemoteDevOps / SRE

Build high-scale observability pipelines and alerting engines handling 1M+ RPS for logs/metrics, develop Golang/Rust gRPC services and APIs, and manage immutable infrastructure with Terraform/Ansible in a distributed systems environment.

Lightning AI

Lightning AI

New York, NY
Infrastructure Engineer (GPU & Compute)
$180k+/yrRemote5+ YOEDevOps / SRE

Owns GPU diagnostics, validation workflows, and automation for bare-metal infrastructure supporting AI/ML workloads. Requires 5+ years in systems engineering with strong Linux, Python, and NVIDIA tools expertise.