Skip to content
296 jobs

Job results

Pindrop

Pindrop

Ukraine

Senior Engineer, Cloud Operations
No salary listedRemote5+ YOEDevOps / SRE

The Senior Cloud Operations Engineer designs, operates, and improves highly available multi-cloud infrastructure, observability, automation, and distributed systems across AWS and GCP. The role requires 5–7 years of DevOps or SRE experience, strong Kubernetes and infrastructure-as-code expertise, and participation in on-call operations.

Mntn

Mntn

Austin, TX

Senior DevOps Engineer
No salary listedRemote5+ YOEDevOps / SRE

Senior security-focused DevOps engineer responsible for improving platform security through architecture, automation, secure defaults, and observability. The role partners with engineering teams across Google Cloud, GKE, Kubernetes, identity, networking, CI/CD, and software supply chain security.

PostHog

PostHog

Remote

ClickHouse Operations Engineer
No salary listedRemoteDevOps / SRE

Automate, manage, and optimize large-scale ClickHouse clusters handling trillions of events and 100+ PB data. Build provisioning systems with Terraform, Ansible, Kubernetes; focus on performance, scaling, and bleeding-edge features.

Airbnb

Airbnb

United States

Staff Software Engineer, Service Tools
$212k+/yrRemote9+ YOEDevOps / SRE

Leads technical direction for Airbnb’s service developer tooling platform, spanning AI-assisted development, JVM build infrastructure, testing, modernization, and observability. Requires 9+ years of industry experience, strong backend and distributed-systems expertise, and the ability to influence organizations and deliver multi-quarter infrastructure initiatives.

Worth AI

Worth AI

Orlando, FL

Senior DevOps Engineer, Infrastructure & Reliability
No salary listedRemote8+ YOEDevOps / SRE

Build and operate reliable, secure cloud infrastructure across AWS and Kubernetes while automating delivery, observability, disaster recovery, and cost optimization. The role requires 8+ years in DevOps, SRE, or infrastructure engineering and strong hands-on experience with Terraform, Kubernetes, AWS, and CI/CD.

Mozilla

Mozilla

Canada

Senior Staff Performance Engineer, Firefox
CA$149k+/yrRemote7+ YOEDevOps / SRE

Leads Firefox performance engineering by writing code, profiling bottlenecks, improving benchmarks, and guiding cross-functional teams. Requires 7+ years of experience, strong C++ and JavaScript skills, and expertise in performance-critical software, profiling, concurrency, and systems analysis.

Webflow

Webflow

United States

Senior Platform Engineer, Infrastructure
$187k+/yrRemote5+ YOEDevOps / SRE

Build and own Webflow’s corporate cloud foundation, including landing zones, networking, security, Infrastructure as Code, GitHub delivery pipelines, self-service deployment patterns, and observability. The role requires 5+ years of platform or cloud engineering experience and strong AWS, Azure, or GCP expertise.

Dataiku

Dataiku

Paris, France

Senior Infrastructure Engineer
No salary listedRemote5+ YOEDevOps / SRE

The Senior Infrastructure Engineer designs and operates internal data platforms and production web-service environments, develops cloud and Linux integrations, and ensures capacity and security. The role requires strong Terraform, Kubernetes, Python, Linux, networking, and cloud-provider experience.

Mercury

Mercury

San Francisco, CA
Software Engineer - Infrastructure
$116k+/yrRemote2+ YOEDevOps / SRE

Build Mercury’s secure, observable infrastructure platform across AWS, networking, containers, and developer tooling. The role requires strong Linux fundamentals, cloud-native experience, technical writing ability, and software development skills, with opportunities to support AI-agent infrastructure.

Coinbase

Coinbase

United States

Staff Infrastructure Engineer, Trading
$218k+/yrRemote8+ YOEDevOps / SRE

Own the infrastructure, deployment, and operational tooling for Coinbase’s latency-sensitive institutional trading platform across cloud and colocated environments. The role requires 8+ years of infrastructure, platform, or SRE experience, strong Linux and networking fundamentals, and experience operating regulated, low-latency systems.

Lightspark

Lightspark

Remote

Senior Production Engineer
$200k+/yrRemote5+ YOEDevOps / SRE

The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.

Komodo Health

Komodo Health

United States

Staff Infrastructure Engineer
$187k+/yrRemote8+ YOEDevOps / SRE

Leads architecture, ownership, modernization, and operation of Komodo Health’s AWS and Kubernetes infrastructure and shared services. The role requires 8+ years of infrastructure experience, deep Terraform and Kubernetes expertise, regulated-environment security fluency, and the ability to establish AI-assisted engineering standards.

Ipfabric

Ipfabric

Prague, Czechia

DevOps Engineer - Platform
No salary listedRemoteDevOps / SRE

Hands-on DevOps platform engineer building product-platform tooling, containerized deployments, CI/CD, and developer-experience improvements. The role requires strong Linux, containers, Kubernetes, troubleshooting, and Python or TypeScript development skills, with L2 and L3 ownership scopes available.

Reltio

Reltio

Bengaluru, India

Senior Engineer
No salary listedRemote6+ YOEDevOps / SRE

Build and operate reliability and resilience capabilities for a multi-cloud platform, including chaos engineering, observability-driven validation, failover testing, and resilient distributed systems. The role requires 6–9 years of software engineering experience and hands-on expertise with Java, Kubernetes, cloud platforms, and CI/CD.

Temporal

Temporal

United States

Staff Software Engineer, Traffic
$212k+/yrRemote8+ YOEDevOps / SRE

Leads the design and development of scalable, secure network traffic systems and cloud infrastructure. The role requires 8+ years of coding experience, strong distributed-systems and concurrency expertise, and deep knowledge of networking and performance optimization.

Imply

Imply

United States

Senior Software Engineer
$155k+/yrRemote6+ YOEDevOps / SRE

Build and operate highly available, distributed platform services and cloud infrastructure for petabyte-scale observability products. The role requires 6+ years of experience, strong Java and AWS expertise, Kubernetes and Terraform production experience, and a bachelor’s degree or equivalent.

Reddit

Reddit

United States

Staff Software Engineer, Observability
$217k+/yrRemote7+ YOEDevOps / SRE

Build and operate Reddit’s internet-scale observability platform across monitoring, logging, and distributed tracing. The role requires 7+ years of infrastructure or software engineering experience, distributed systems expertise, and strong Kubernetes and troubleshooting skills.

Fal

Fal

Remote

Senior/Staff Kubernetes Infrastructure Engineer
$180k+/yrRemote5+ YOEDevOps / SRE

Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.

Phantom

Phantom

Remote

Staff DevOps Engineer
No salary listedRemote8+ YOEDevOps / SRE

Owns and evolves CI/CD, mobile release, testing, and deployment infrastructure for a production fintech application. The role requires 8+ years in DevOps or related platform disciplines, strong AWS and Kubernetes expertise, and experience with secure mobile release systems.

Grafana Labs

Grafana Labs

United Kingdom
Software Engineer - Platform Metal
£72k+/yrRemoteDevOps / SRE

Build and operate Grafana’s physical infrastructure platform, including bare-metal environments, Kubernetes clusters, networking, scheduling, and autoscaling. The role requires datacenter and software-operations experience, with strong skills in Kubernetes and infrastructure automation using tools such as Go, Terraform, and Crossplane.

ZoomInfo

ZoomInfo

Toronto, Canada

Senior Software Engineer - CI/CD
No salary listedRemote5+ YOEDevOps / SRE

Builds and operates CI/CD platforms across GitHub Actions, Jenkins, Kubernetes, and cloud infrastructure. The role owns GitOps delivery, reusable developer tooling, observability, reliability, and migration initiatives across distributed engineering teams.

Twilio

Twilio

Ireland

DevOps Engineer
No salary listedRemoteDevOps / SRE

Build and lead the evolution of Twilio’s large-scale observability platform, including telemetry pipelines, query systems, developer tooling, and standards. The role requires expertise in observability systems, distributed systems, cloud infrastructure, and modern programming languages.

Twilio

Twilio

India

Senior Network Engineer
No salary listedRemote7+ YOEDevOps / SRE

Builds and operates Twilio’s global corporate network, VPN, zero-trust access, and cloud connectivity while monitoring performance and resolving incidents. The role requires substantial experience with Cisco, Palo Alto Networks, AWS networking, security protocols, and enterprise troubleshooting.

Skydio

Skydio

San Mateo, CA
Staff Site Reliability Engineer
$240k+/yrRemote8+ YOEDevOps / SRE

Owns and scales production cloud infrastructure across Kubernetes/EKS, AWS, Terraform, CI/CD, networking, and observability. The role requires 8+ years of infrastructure experience, strong Kubernetes operations expertise, and depth in reliability or scaling challenges.

Attentive

Attentive

United States

Staff Site Reliability Engineer
$180k+/yrRemote7+ YOEDevOps / SRE

Leads strategic production engineering initiatives that improve the reliability, scalability, observability, and security of large-scale platforms. The role requires 7+ years of relevant experience, strong coding skills, and expertise in reliability practices such as SLIs, SLOs, and incident management.

Clickhouse

Clickhouse

Remote

Senior Cloud Software Engineer - Efficiency Engineering
No salary listedRemote5+ YOEDevOps / SRE

Designs and operates scalable, highly available cloud infrastructure while leading efficiency initiatives across compute, storage, networking, and cost optimization. Requires 5+ years of distributed-systems software development experience and expertise with cloud platforms, infrastructure as code, and Kubernetes.

Clickhouse

Clickhouse

Remote

Senior Cloud Software Engineer - Efficiency Engineering
No salary listedRemote5+ YOEDevOps / SRE

Build and optimize ClickHouse Cloud’s highly available, multi-cloud infrastructure, including automation, distributed systems, networking, security, and cost-efficiency tooling. Requires 5+ years of experience operating scalable systems and expertise in cloud platforms, infrastructure as code, and production engineering.

Temporal

Temporal

United States

Senior Software Engineer, Infrastructure Foundations
$176k+/yrRemote10+ YOEDevOps / SRE

Build and scale reliable cloud infrastructure systems, shape long-term architecture and roadmaps, and drive cross-functional alignment. The role requires 10+ years of coding experience, distributed-systems and concurrency expertise, deep infrastructure experience, and hands-on cloud-provider experience.

Fluidstack

Fluidstack

Remote

Principal Operations Engineer, Mechanical
$150k+/yrRemote10+ YOEDevOps / SRE

As a Principal Operations Engineer, Mechanical, you will be the senior technical authority for mechanical and cooling infrastructure across hyperscale AI data centers. You will lead site assessments, drive operational readiness, review designs, and ensure precision execution of critical systems.

Alpaca

Alpaca

Americas
Senior DevOps Engineer
No salary listedRemote5+ YOEDevOps / SRE

Designs and operates highly available GCP infrastructure and developer platforms for trading-critical systems. The role requires 5+ years of DevOps, platform, infrastructure, or SRE experience, with strong Terraform, Kubernetes, networking, CI/CD, observability, and incident-management skills.

Orkes

Orkes

EMEA

Site Reliability Engineer
$125k+/yrRemote5+ YOEDevOps / SRE

Owns reliability, observability, incident response, and automation for cloud-based production systems. The role requires 5+ years in SRE, DevOps, platform engineering, or related infrastructure work, with strong Kubernetes, cloud, distributed-systems, and infrastructure-automation experience.

OpenSea

OpenSea

United States

Staff Platform Engineer
$190k+/yrRemote7+ YOEDevOps / SRE

Build and operate scalable platform services, infrastructure, and developer tooling that enable reliable product delivery. The role requires 7+ years of software engineering experience, JVM expertise, distributed-systems experience, and strong platform, cloud, CI/CD, and observability skills.

Headway

Headway

San Francisco, CA
Staff Infrastructure Engineer
$265k+/yrRemote8+ YOEDevOps / SRE

Own the cloud platform, deployment architecture, container infrastructure, networking, autoscaling, cost controls, and Python runtime health for a high-scale healthcare technology platform. The role requires 8+ years in infrastructure, platform, or SRE work, deep AWS expertise, Terraform experience, and Staff-level cross-team influence.

Prove AI

Prove AI

Ireland

Senior Manager, Platform Engineering
€120k+/yrRemote5+ YOEDevOps / SRE

Leads Ireland-based Platform Developer Enablement and SRE teams, defining platform strategy, developer self-service, reliability objectives, and observability standards. Requires senior software, SRE, or platform engineering experience, management leadership, and expertise in cloud infrastructure, Kubernetes, Terraform, CI/CD, and distributed systems.

MongoDB

MongoDB

Cork, Ireland
Site Reliability Engineer , Storage Layer Services
No salary listedRemote6+ YOEDevOps / SRE

This SRE will operate and improve MongoDB Atlas’s multi-tenant distributed storage infrastructure, focusing on reliability, performance, observability, automation, and incident response. The role requires 6+ years of distributed-systems experience plus expertise in storage or databases, Kubernetes, cloud platforms, Linux, and networking.

Kraken

Kraken

LATAM

Site Reliability Engineer - Telemetry
No salary listedRemote3+ YOEDevOps / SRE

Operates and scales shared telemetry infrastructure spanning metrics, logs, traces, alerting, dashboards, and profiling. The role requires at least three years of production engineering experience, distributed-systems troubleshooting, Infrastructure as Code, container orchestration, incident response, and on-call participation.

Shield AI

Shield AI

United States

Sr. Staff Platform/Data Reliability Engineer, Databricks
$180k+/yrRemote12+ YOEDevOps / SRE

Leads the operational reliability, security, observability, deployment standards, and governance of Databricks for enterprise data workloads. Requires 12+ years in platform, SRE, or cloud data infrastructure engineering plus production Databricks experience and expertise in CI/CD, secure execution, and regulated environments.

Tatari

Tatari

Poland
Senior SRE
No salary listedRemote5+ YOEDevOps / SRE

The Senior SRE will design, automate, and operate high-throughput AWS and Kubernetes infrastructure, improving reliability, observability, CI/CD, and cost efficiency. The role requires 4–6 years of production SRE, DevOps, or systems engineering experience and strong Terraform, Linux, scripting, and incident-response skills.

GitLab

GitLab

Canada
Site Reliability Engineer, Intermediate to Senior Staff
$126k+/yrRemote5+ YOEDevOps / SRE

Site Reliability Engineers build and operate scalable production infrastructure, automate operational workflows, and improve observability, incident response, and service reliability. The role spans Intermediate through Senior Staff levels and requires experience with Kubernetes, infrastructure as code, cloud platforms, and software engineering.

Grafana Labs

Grafana Labs

United Kingdom
Staff Software Engineer - Databases SRE
£104k+/yrRemote8+ YOEDevOps / SRE

Leads production reliability for Grafana Cloud’s multi-tenant database products, partnering with product engineering teams to improve SLOs, scalability, observability, automation, and incident response. Requires 8+ years of engineering experience, including substantial SRE or production engineering work, plus strong Kubernetes and cloud expertise.

OnePay

OnePay

United States

Platform Engineer
$170k+/yrRemote7+ YOEDevOps / SRE

Builds and operates core platform services, Kafka-based event streaming, and developer frameworks for high-scale distributed systems in fintech. Requires 7+ years experience with AWS, Kubernetes, and cloud-native infrastructure.

Mattermost

Mattermost

United States

Lead Site Reliability Engineer
$145k+/yrRemote7+ YOEDevOps / SRE

Leads the architecture, reliability, observability, and operational excellence of secure cloud and hybrid infrastructure for a mission-critical collaboration platform. Requires 5+ years in SRE, DevOps, or cloud infrastructure, with expertise in Kubernetes, Terraform, AWS, and regulated environments.

Solace

Solace

United States

Senior Platform Engineer
No salary listedRemote5+ YOEDevOps / SRE

Build and operate scalable cloud infrastructure and application platforms, enabling frequent deployments, resilient systems, observability, and self-healing capabilities. The role requires strong troubleshooting, Linux and cloud experience, networking knowledge, and expertise in one or more platform engineering focus areas.

Kraken

Kraken

LATAM

Infrastructure Engineer - Core Infrastructure
No salary listedRemote3+ YOEDevOps / SRE

Operates and scales Kraken’s core infrastructure platforms, with a focus on OpenStack, Ceph, Linux, distributed systems, and automation. The role requires 3+ years of infrastructure or software engineering experience and supports reliable compute and storage services across cloud and on-premises environments.

Webflow

Webflow

Argentina

Senior Infrastructure Engineer
No salary listedRemote5+ YOEDevOps / SRE

Owns and evolves Webflow’s highly available, multi-cloud infrastructure, including Kubernetes, networking, infrastructure as code, observability, and AI-powered automation. The role requires 5+ years operating customer-facing cloud infrastructure and deep AWS experience.

Webflow

Webflow

Argentina

Staff DevOps Engineer, Delivery Loop
No salary listedRemote7+ YOEDevOps / SRE

Leads Webflow’s deployment strategy and GitOps platform, improving CI/CD reliability, progressive delivery, and developer productivity. The role requires 7+ years in DevOps, SRE, or infrastructure engineering and deep experience with Kubernetes, AWS, Docker, and infrastructure as code.

PostHog

PostHog

United States

Site Reliability Engineer
No salary listedRemote5+ YOEDevOps / SRE

Site Reliability Engineer responsible for operating and scaling a large multi-region, multi-account AWS + Kubernetes platform. Focus on automation, IaC with Terraform/Terragrunt, reducing operational toil, and owning production stateful systems end-to-end including on-call.

PostHog

PostHog

San Francisco, CA

SRE - Infra
No salary listedRemoteDevOps / SRE

Owns and automates production infrastructure on multi-region AWS with EKS clusters, focusing on scaling, reliability, and self-healing systems. Requires deep Kubernetes, Terraform, and Linux expertise for large-scale stateful workloads.

Dropbox

Dropbox

Poland

Infrastructure Software Engineer
PLN272k+/yrRemote5+ YOEDevOps / SRE

Build and operate large-scale, geographically distributed infrastructure supporting massive file metadata, data volumes, analytics, and concurrent connections. The role requires 5+ years of software development experience and expertise in backend systems, programming, operating systems, and distributed infrastructure.

Alpaca

Alpaca

United States

Senior Site Reliability Engineer
No salary listedRemote5+ YOEDevOps / SRE

Operates and improves reliability for a trading-critical brokerage platform across cloud infrastructure, Kubernetes, observability, messaging, and PostgreSQL. Requires 4+ years of production operations experience, strong PostgreSQL fundamentals, incident response expertise, and proficiency in Go or Python.