Software Engineer: Infrastructure
Infrastructure Engineer owns platform reliability, security, and performance in a HIPAA-compliant environment handling healthcare data. Designs infrastructure with Terraform, improves CI/CD workflows, and ensures resilient systems using PostgreSQL and modern cloud tools.
About the job
Responsibilities
- Own infrastructure health, performance, and reliability across the stack.
- Make architectural decisions and improve how we build and operate.
- Design and evolve infrastructure using Terraform.
- Improve CI/CD, deployments, and developer workflows so teams ship faster.
- Strengthen security across access controls, vulnerability management, and dependencies.
- Identify and eliminate reliability risks before they impact customers.
Requirements
- Experience operating production infrastructure, platform, DevOps, or SRE systems.
- A track record of owning critical systems end-to-end.
- Strong experience with PostgreSQL or comparable relational databases in production.
- Experience managing infrastructure as code using Terraform.
- Sound judgment across reliability, performance, and security tradeoffs.
- A bias toward automation and improving developer productivity.
Nice-to-Haves
- Experience operating in HIPAA or SOC 2 environments.
- Deep PostgreSQL performance tuning or replication management.
- Building internal tooling that improves developer productivity.
- Improving observability, incident response, or overall security posture.
Tech Stack
- Ruby on Rails
- PostgreSQL
- Terraform
- Modern cloud infrastructure and CI/CD tooling
Compensation
$180,000—$250,000 USD
Skills
Terraform, Postgres, Ruby on Rails, CI/CD, DevOps, SRE, HIPAA, SOC 2, Infrastructure As Code, Cloud Infrastructure
Similar jobs
DevOps / SRE jobsBuild developer-experience tooling and release systems within Benchling’s Platform team, helping engineering teams develop, test, package, and ship high-quality software rapidly. The role requires 4+ years of software engineering experience, web framework expertise, strong problem-solving, and effective cross-functional communication.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Own Mercor’s internal identity and cloud platform infrastructure as code, automating provisioning, access management, secrets, and employee lifecycle workflows. The role requires production Terraform, Okta, SCIM, and multi-cloud IAM experience, plus strong automation, incident response, and documentation skills.
Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.
Infrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.