Skip to content

Cloud Infrastructure Engineer

Builds and maintains scalable cloud infrastructure using Terraform and Kubernetes, owns CI/CD pipelines and observability, supports multi-cloud deployments, and assists customers with self-hosting. Requires 5+ years in DevOps/SRE, deep AWS experience, and programming skills.

About the job

What you’ll do

  • Build and maintain Terraform modules for both internal infrastructure and customer deployments
  • Work directly with customers in Slack to support self-hosting and troubleshoot infrastructure issues. Build tools to make it easier for them to support themselves.
  • Own and improve our CI/CD pipeline: reduce build times, improve failure visibility, and enable safer, faster releases
  • Centralize and scale observability - including logs, metrics, dashboards, and alerts
  • Partner with engineering teams to build and evolve a secure, developer-friendly infrastructure platform
  • Support multi-cloud deployment patterns (AWS primarily, with Azure and GCP support for enterprise customers)
  • Implement tools and automation to improve deployment, rollback, and infrastructure reliability

Ideal candidate credentials

  • 5+ years of experience in DevOps, SRE, or Infrastructure Engineering roles
  • Deep experience with Terraform and at least one major cloud provider (AWS strongly preferred)
  • Strong Kubernetes skills: deploying, debugging, and scaling real workloads
  • Proficient in scripting or programming (Python, Typescript, or Go)
  • Experience supporting production systems and responding to incidents
  • Comfortable working directly with customers in a support or deployment context

Bonus: experience with multi-cloud environments or self-hosted enterprise software

Benefits

  • Medical, dental, and vision insurance
  • Daily lunch, snacks, and beverages
  • Flexible time off
  • Competitive salary and equity
  • AI Stipend

Skills

Terraform, Kubernetes, AWS, CI/CD, Observability, Python, TypeScript, Go, Azure, GCP

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.