Senior Software Engineer building Chainguard's internal Developer Platform "Factory", including monorepo CI/CD pipelines, Agentic AI platform for automated changes, and paved-road build infrastructure to reduce developer toil and accelerate secure artifact delivery.
157k – 184k/yr
Remote4+ YOEDevOps / SRE
About the role
What you'll do
Collaborate on key areas of the Developer Platform contributing to technical implementation alongside the team.
Help improve the speed, reliability, and developer experience of the monorepo CI/CD pipeline so engineers spend more time shipping and less time waiting on or debugging builds.
Help productionize the Agentic AI platform: agent observability, RAG, and the context-engineering patterns that keep agents reliable from dev through prod.
Partner with the team to consolidate fragmented build systems and deliver standardized “paved road” blueprints so new products can onboard and ship to production quickly.
Partner with engineers across the company to identify toil, codify patterns, and automate repetitive manual work.
Participate in design and code reviews, and contribute your technical voice to engineering velocity and operational excellence.
What we're looking for
4+ years of experience building and operating production services and platform infrastructure in modern cloud environments.
Proficiency with Go (Golang), or strong readiness to ramp quickly.
Experience with CI/CD systems, container-based orchestration, and the tooling required to scale a large monorepo across build, test, and release.
A strong AI-forward mindset; ability to connect your past technical experience with modern, AI-driven development practices (e.g., LLM-based automation, context engineering).
Excellent communication and collaboration skills; able to work with engineers across the company to identify and automate away repetitive manual tasks.
A genuine fit with a high-paced, “ship it” culture: self-directed, comfortable with ambiguity, and focused on data-driven improvements to product delivery.
Builds and operates scalable, reliable infrastructure for high-traffic AI platform using cloud providers, containers, and IaC tools. Owns observability, SLOs, incident response, and collaborates on system design. Requires 7+ years in infrastructure/DevOps with strong programming skills.
158k – 278k/yr
Hybrid7+ YOEDevOps / SRE
Senior Manager, DevOps
PindropUnited States
Lead DevOps strategy and team to improve engineering velocity, platform reliability, and operational efficiency across multi-cloud (AWS/GCP) environments. Drive IaC, Kubernetes delivery, observability, AI-powered tooling adoption, and cross-functional collaboration.
155k – 185k/yr
Remote6+ YOEDevOps / SRE
Senior Software Engineer, AI Native Web Platform
DatabricksMountain View, CA
Senior Software Engineer building and scaling the AI-native web platform layer at Databricks, including CI/CD, deployment automation, consent management, accessibility testing, and tech stack unification. Requires 8+ years experience in production web infrastructure.
160k – 220k/yr
On-site8+ YOEDevOps / SRE
Senior Site Reliability Engineer
TalkiatryUnited States
Join as the first SRE to define reliability practices, SLOs, observability, and toil reduction across six product teams at a leading mental health platform. 7+ years software/infra engineering with hands-on SRE experience required; product teams retain on-call ownership.
160k – 185k/yr
Remote7+ YOEDevOps / SRE
Senior Infrastructure Engineer
AurelianSeattle, WA
Senior Infrastructure Engineer building analytics, observability, and developer tooling for Aurelian's real-time AI agents used in 911 emergency response centers. Requires 4+ years in infrastructure/platform/backend roles with experience in reliability and scale.