Skip to content
OpenlyOpenlyUnited States

Site Reliability Engineer II (Remote, US)

Build and maintain infrastructure for Openly's insurance platform as a DevOps/SRE II. Focus on automation, IaC with Terraform, monitoring, incident response, and reducing toil on Google Cloud with Kubernetes. Requires 2+ years infrastructure automation experience, strong Python/Go scripting, and cloud expertise.

115k – 173k/yr
Remote2+ YOEDevOps / SRE

About the role

Key Responsibilities

  • Build internal tooling to help other engineers and the rest of the company understand and operate our system
  • Design and implement security best practices for our team and infrastructure
  • Reduce toil through automation, including building and maintaining CI/CD infrastructure
  • Build infrastructure as code using declarative provisioning tools
  • Develop high signal-to-noise ratio monitoring and alerting policies and technology to help us meet our SLOs
  • Lead incident response and postmortems
  • Contribute to important architectural and operational decisions like microservices vs. monoliths, deployment techniques, technologies, policies, etc.

Requirements

  • 2+ years of professional/production experience developing and using infrastructure automation tools and techniques
  • Proven track record of creating improvements in business-critical systems around stability, performance, and scalability
  • Demonstrated ability to deliver complete systems from start to finish in a reasonable time frame
  • Understands the consequences of running software in production and are willing to share your knowledge with the rest of the team
  • Ability to explain complex technical challenges to non-technical audiences
  • Strong scripting skills in one or more of the following: Python, Go
  • Experience working with Infrastructure as Code (IaC) tooling, preferably Terraform
  • Cloud experience

Nice-to-Haves

  • Experience with our stack: Google Cloud (Cloud Run, Kubernetes, Pub/Sub, BigQuery, CloudSQL), Terraform, GitHub, DataDog, CircleCI
  • Backend: Go & PostgreSQL
  • Data: GCP GCS, BigQuery, Composer/Airflow, Cloud Functions, Postgres, SQL, Python, Go, Aiven Debezium and Kafka, Fivetran
  • Frontend exposure: VueJS, Webpack, Nuxt, Tailwind

Compensation & Benefits

Budgeted Salary Range: $115,200 - $129,600 USD
Full Salary Range: $115,200 - $172,800 USD

  • Remote-First Culture
  • Competitive Salary & Equity
  • Comprehensive Medical, Dental, and Vision Plan Offerings
  • Life and disability coverage
  • Parental Leave - up to 8 weeks
  • 401K Company Contribution (3% of gross income)
  • Work-from-home stipend ($1,500)
  • Annual Professional Development Fund ($2,000)
  • Be Well Program ($50/month)
  • Paid Volunteer Service Hours
  • Referral Program

Skills

TerraformKubernetesGCPPythonGoCI/CDDatadogCircleCIPostgresBigQueryInfrastructure As CodeMonitoringSRE

Similar roles

DevOps / SRE jobs
Twilio

Software Engineer L2

TwilioNew York, NY +3

Software Engineer L2 responsible for evolving and maintaining Twilio's Compute infrastructure, including VM orchestration, AWS Auto Scaling Groups, hardened AMIs, secure container images, and automation of operational tasks in a remote-first environment.

117k – 172k/yr
Remote2+ YOEDevOps / SRE
Applied Intuition

Software Engineer - Developer Infrastructure

Applied IntuitionSunnyvale, CA

Builds and improves core libraries, frameworks, and developer tools like Bazel and Buildkite CI/CD to boost engineering productivity. Requires 2+ years experience, Bachelor's in CS, and expertise in Go/C++/Python/TypeScript.

120k – 300k/yr
On-site2+ YOEDevOps / SRE
The Voleon Group

Site Reliability Engineer

The Voleon GroupNew York, NY +1

Site Reliability Engineer improves, manages, and monitors production-critical infrastructure and data pipelines in a finance AI/ML firm. Collaborates on fault-tolerance, deployments, automation, and on-call incident response using Python, Linux, and cloud tools. Requires 2+ years experience and quantitative degree.

120k – 160k/yr
Remote2+ YOEDevOps / SRE
Baseten

Capacity Ops Associate

BasetenSan Francisco, CA +1

Manages GPU fleet operations, including node maintenance, capacity fulfillment, and technical orchestration between SRE/infra teams and customers. Requires 2+ years experience, Kubernetes familiarity, and strong communication skills.

120k – 160k/yr
Hybrid2+ YOEDevOps / SRE
EliseAI

Platform Operations Engineer

EliseAINew York, NY

Leads cross-functional technical projects to optimize tech stack, build custom automation and analytics solutions for business operations, and integrate systems using AWS, APIs, and databases. Requires 2+ years experience with Python/SQL proficiency and onsite presence in New York.

120k – 200k/yr
On-site2+ YOEDevOps / SRE