Skip to content
TennrTennr

Senior Infrastructure Engineer

Own and scale Tennr’s AWS infrastructure, Kubernetes environments, and infrastructure-as-code foundation across development, staging, and production. The role requires 5–8 years of infrastructure, platform, or DevOps experience, strong Kubernetes and AWS expertise, and hands-on production ownership.

About the job

Responsibilities

  • Build the EKS cluster module and reusable infrastructure patterns.
  • Own cluster management and end-to-end environments for development, staging, and production.
  • Drive and unblock infrastructure migrations using safe, staged approaches.
  • Consolidate infrastructure observability onto Datadog and establish operational standards.
  • Build and maintain AWS infrastructure, including networking, IAM, and core services, using infrastructure as code.
  • Keep infrastructure secure by default, reliable, and cost-aware.

Requirements

  • 5–8 years of experience in infrastructure, platform, or DevOps engineering with ownership of production systems.
  • Hands-on experience with Kubernetes, preferably Amazon EKS.
  • Experience with Terraform or comparable infrastructure-as-code tools.
  • Strong AWS fundamentals, including networking, IAM, and core cloud services.
  • Experience with Datadog or a comparable observability stack.
  • Ability to execute autonomously and make pragmatic infrastructure decisions.
  • Bias toward simple, durable systems and shipping over over-engineering.

Nice-to-haves

  • Experience working in regulated or healthcare environments.
  • Familiarity with HIPAA and protected health information (PHI).

Compensation and Benefits

  • Annual salary: $200,000–$230,000.
  • Unlimited PTO.
  • 100% paid employee health benefit options.
  • Employer-funded 401(k) match.
  • Competitive parental leave.
  • Free lunch and snacks.

Skills

Kubernetes, Amazon Eks, Terraform, AWS, Aws Networking, Aws Iam, Datadog, Infrastructure As Code, Observability, Cloud Infrastructure

Skydio

Skydio

San Mateo, CA

Senior Software Engineer, Developer Productivity
$200k+/yrOn-site5+ YOEDevOps / SRE

Build and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.

Anyscale

Anyscale

San Francisco, CA

Senior Site Reliability Engineer, Platform Infrastructure
$200k+/yrHybrid5+ YOEDevOps / SRE

Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.

Onos Health

Onos Health

San Francisco, CA

Lead Infrastructure Engineer
$200k+/yrHybrid7+ YOEDevOps / SRE

Leads infrastructure and platform strategy for a production healthcare AI platform, owning AWS, reliability, disaster recovery, compliance, CI/CD, and secure AI-agent operations. Requires deep cloud and Terraform expertise, audit-cycle experience, and prior technical leadership.

Lightspark

Lightspark

Remote

Senior Production Engineer
$200k+/yrRemote5+ YOEDevOps / SRE

The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.

Garner Health

Garner Health

United States

Senior Site Reliability Engineer
$191k+/yrRemote5+ YOEDevOps / SRE

Own the reliability, resilience, observability, and automation of AWS and Kubernetes infrastructure supporting production products and AI/ML workloads. The role requires 4+ years of cloud infrastructure experience, strong Kubernetes and Terraform expertise, and senior-level incident response and software engineering skills.