Skip to content
ThatchThatch

Software Engineer: Infrastructure

Infrastructure Engineer owns platform reliability, security, and performance in a HIPAA-compliant environment handling healthcare data. Designs infrastructure with Terraform, improves CI/CD workflows, and ensures resilient systems using PostgreSQL and modern cloud tools.

About the job

Responsibilities

  • Own infrastructure health, performance, and reliability across the stack.
  • Make architectural decisions and improve how we build and operate.
  • Design and evolve infrastructure using Terraform.
  • Improve CI/CD, deployments, and developer workflows so teams ship faster.
  • Strengthen security across access controls, vulnerability management, and dependencies.
  • Identify and eliminate reliability risks before they impact customers.

Requirements

  • Experience operating production infrastructure, platform, DevOps, or SRE systems.
  • A track record of owning critical systems end-to-end.
  • Strong experience with PostgreSQL or comparable relational databases in production.
  • Experience managing infrastructure as code using Terraform.
  • Sound judgment across reliability, performance, and security tradeoffs.
  • A bias toward automation and improving developer productivity.

Nice-to-Haves

  • Experience operating in HIPAA or SOC 2 environments.
  • Deep PostgreSQL performance tuning or replication management.
  • Building internal tooling that improves developer productivity.
  • Improving observability, incident response, or overall security posture.

Tech Stack

  • Ruby on Rails
  • PostgreSQL
  • Terraform
  • Modern cloud infrastructure and CI/CD tooling

Compensation

$180,000—$250,000 USD

Skills

Terraform, Postgres, Ruby on Rails, CI/CD, DevOps, SRE, HIPAA, SOC 2, Infrastructure As Code, Cloud Infrastructure

Benchling

Benchling

San Francisco, CA
Software Engineer, Platform
$173k+/yrHybrid4+ YOEDevOps / SRE

Build developer-experience tooling and release systems within Benchling’s Platform team, helping engineering teams develop, test, package, and ship high-quality software rapidly. The role requires 4+ years of software engineering experience, web framework expertise, strong problem-solving, and effective cross-functional communication.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Mercor

Mercor

San Francisco, CA

Cloud Platform Engineer
$190k+/yrOn-siteDevOps / SRE

Own Mercor’s internal identity and cloud platform infrastructure as code, automating provisioning, access management, secrets, and employee lifecycle workflows. The role requires production Terraform, Okta, SCIM, and multi-cloud IAM experience, plus strong automation, incident response, and documentation skills.

Ramp

Ramp

New York, NY
TLM, Production Engineering
$168k+/yrHybrid3+ YOEDevOps / SRE

Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.

Roboflow

Roboflow

New York, NY
Infrastructure Engineer
$165k+/yrRemoteDevOps / SRE

Infrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.