Skip to content
ForusForus

Software Engineer, Platform & Infrastructure

Software Engineer owning Forus' compute platform (EKS/Kubernetes), data layer (Postgres, OpenSearch, BigQuery), cloud cost optimization, reliability (SLOs, observability), and IaC primitives in a regulated healthcare environment. Requires production Kubernetes at scale, deep AWS/Terraform expertise, and database migration experience.

About the job

Responsibilities

  • Own our compute platform, including container orchestration, networking, autoscaling, service topology, and ingress, and define the architecture and patterns that keep it reliable and simple as we scale.
  • Own our data layer, driving performance, scaling, and reliability across our databases (Postgres, OpenSearch, BigQuery), and lead the platform decisions that keep our datastores fast and cost-effective as load grows.
  • Own cloud spend by building attribution and driving optimization across compute, data, and storage, so spend grows more slowly than the business does.
  • Define and own platform reliability, including SLOs, capacity planning, infrastructure observability, and incident response for the systems every product team depends on.
  • Build the infrastructure-as-code, service patterns, and self-serve primitives that let product teams provision and operate services safely without waiting on you.

Requirements

  • A track record of operating production Kubernetes at scale, including networking, autoscaling, and cluster reliability.
  • Deep AWS expertise (EKS, VPC, IAM boundaries) and strong infrastructure-as-code with Terraform. You build platforms and reusable primitives rather than one-off resources.
  • Run a real, low-downtime database migration and operated a relational store like Postgres under production load, with the ability to reason from a query plan to a capacity plan.
  • Strong instincts for reducing complexity. You are drawn to consolidating and simplifying systems rather than adding surface area.
  • Comfort operating in a regulated, high-stakes environment. Familiarity with PHI, HIPAA, or SOC 2 is a strong plus, as is a track record of mastering a high-stakes domain quickly.
  • A track record of moving quickly, finding shortcuts, and going to unreasonable lengths to deliver on goals.
  • High NPS with your former teammates.

Nice-to-Haves

  • Familiarity with PHI, HIPAA, or SOC 2.

Compensation and Benefits

  • Competitive compensation with meaningful equity.
  • Fully covered medical, vision, and dental insurance.
  • Memberships for One Medical, Talkspace, Teladoc, and Kindbody.
  • Unlimited paid time off (PTO) and 16 weeks of parental leave.
  • 401K plan setup, FSA option, commuter benefits, and DashPass.
  • Lunch at the office every day and Dinner at the office after 7 pm.
  • Salary ranges based on paying competitively for company size and industry; individual pay decisions based on qualifications, experience, skillset, geography, and internal equity.

Skills

Kubernetes, AWS, EKS, Terraform, Postgres, Opensearch, BigQuery, Vpc, IAM

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.