Build and operate internal developer platforms, cloud infrastructure, CI/CD systems, and observability tooling that improve engineering productivity and production reliability. The role requires hands-on experience with AWS, Kubernetes, distributed systems, and infrastructure automation.
160k – 250k/yr
On-site5+ YOEDevOps / SRE
About the role
Responsibilities
Developer Productivity & Internal Platform
Build and evolve the internal developer platform (IDP).
Create self-service tooling for provisioning services, environments, and infrastructure.
Improve local development workflows, testing, deployment ergonomics, and developer feedback loops.
Reduce operational friction through automation and platform abstractions.
CI/CD & Release Engineering
Design and maintain scalable CI/CD systems using GitHub Actions and infrastructure-as-code workflows.
Improve build times, deployment reliability, rollback, and recovery systems.
Create deployment workflows that are fast, observable, and reliable in production.
Cloud Infrastructure & Kubernetes
Design and operate cloud-native systems on AWS, EKS, Docker, and Kubernetes.
Improve platform scalability, resilience, and developer onboarding.
Define service architecture and deployment standards across the engineering organization.
Observability & Reliability
Own platform observability using Datadog, metrics, logging, and distributed tracing.
Build dashboards, alerts, and operational tooling for engineers.
Improve reliability, performance, and operational visibility across services.
Distributed Systems & Platform Services
Work across core infrastructure systems including PostgreSQL, Kafka, Redis, and Temporal.
Engineering Enablement
Partner with backend and AI engineers to improve platform usability.
Establish engineering standards and platform best practices.
Treat infrastructure as a product rather than a support function.
Nice-to-Haves
Experience building internal developer platforms (IDPs).
Experience in high-growth or early-stage startups.
Leads developer experience and AI tooling across the engineering organization, building internal agents, ephemeral environments, and faster CI/CD workflows. Requires 5+ years in cloud infrastructure, platform, or developer tooling, plus hands-on AI assistant and LLM workflow experience.
160k – 180k/yrHybrid5+ YOEDevOps / SRE
Senior Site Reliability Engineer
TalkiatryUnited States
Join as the first SRE to define reliability practices, SLOs, observability, and toil reduction across six product teams at a leading mental health platform. 7+ years software/infra engineering with hands-on SRE experience required; product teams retain on-call ownership.
160k – 185k/yrRemote7+ YOEDevOps / SRE
Senior Infrastructure Engineer
AurelianSeattle, WA
Senior Infrastructure Engineer building analytics, observability, and developer tooling for Aurelian's real-time AI agents used in 911 emergency response centers. Requires 4+ years in infrastructure/platform/backend roles with experience in reliability and scale.
160k – 220k/yrOn-site4+ YOEDevOps / SRE
Senior Microsoft Cloud Infrastructure Engineer
CrusoeSan Francisco, CA
Senior Cloud Infrastructure Engineer owning design, implementation, and management of Microsoft 365, Entra ID, Azure, and Azure Arc hybrid infrastructure. Requires 8+ years infrastructure experience with deep Azure/M365 expertise, IaC, Windows Server admin, and hands-on data center hardware work.
160k – 195k/yrOn-site8+ YOEDevOps / SRE
Lead DevOps Engineer
OctusNew York, NY
Lead a team of DevOps engineers to design, implement, and maintain CI/CD pipelines, cloud infrastructure, monitoring, and security best practices. Requires 7+ years of DevOps experience including 2 years in leadership.