Skip to content
OpenAIOpenAI

Software Engineer, Delivery / CD

Builds and operates continuous deployment platforms for safe, rapid code rollouts across Kubernetes clusters and global regions. Focuses on progressive delivery, GitOps, automation, and AI-assisted workflows to boost developer productivity.

About the job

In This Role, You Will

  • Design and build continuous deployment infrastructure that safely rolls out changes across dozens of Kubernetes clusters and global regions.
  • Develop systems for progressive delivery, including canary releases, staged rollouts, and automated rollback.
  • Improve engineering velocity by reducing friction in the release pipeline and automating manual operational workflows.
  • Work with product and infrastructure teams to ensure their services are deployable, observable, and resilient at scale.
  • Implement and evolve deployment methodologies such as GitOps, infrastructure-as-code, and progressive delivery patterns.
  • Build systems that automatically evaluate deployment health using metrics, logs, traces, and alerts to detect regressions and trigger safe rollbacks.
  • Build systems that support agent-assisted or autonomous deployment workflows using modern AI tooling.

Technologies commonly used in this environment include:

  • Kubernetes for large-scale container orchestration and runtime infrastructure
  • Python and FastAPI for internal services
  • Terraform for infrastructure as code
  • GitOps-based deployment workflows (e.g., ArgoCD, Flux, or similar systems)
  • Buildkite for CI orchestration

You may be a strong fit if you:

  • Have worked with Kubernetes-based deployment systems at scale
  • Have experience building or operating continuous deployment platforms
  • Are familiar with GitOps tooling such as ArgoCD or Flux
  • Are excited about building AI-assisted systems and agents that intelligently shepherd software changes from commit to safe production rollout.
  • Care deeply about safe production rollouts and minimizing blast radius
  • Enjoy building internal platforms that improve developer productivity across the organization

Compensation

$230K – $490K + Offers Equity

Skills

Kubernetes, Python, FastAPI, Terraform, GitOps, Argo CD, Flux, Buildkite

Perplexity

Perplexity

San Francisco, CA
Member of Technical Staff
$220k+/yrRemote4+ YOEDevOps / SRE

Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.

Tessera Labs

Tessera Labs

San Francisco, CA

AI Platform Engineer
$200k+/yrRemote5+ YOEDevOps / SRE

Build and own production-grade AI agent infrastructure across multiple clouds, with responsibility for Kubernetes, Terraform, observability, security, reliability, and automation. Requires 5+ years of cloud infrastructure experience and strong CI/CD, networking, and production operations expertise.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Mercor

Mercor

San Francisco, CA

Cloud Platform Engineer
$190k+/yrOn-siteDevOps / SRE

Own Mercor’s internal identity and cloud platform infrastructure as code, automating provisioning, access management, secrets, and employee lifecycle workflows. The role requires production Terraform, Okta, SCIM, and multi-cloud IAM experience, plus strong automation, incident response, and documentation skills.

Firecrawl

Firecrawl

San Francisco, CA

Cloud DevOps Engineer
$240k+/yrHybrid5+ YOEDevOps / SRE

Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.