# Senior Software Engineer, Site Reliability

**Company:** [Bloomerang](https://hotfix.jobs/companies/bloomerang)
**Location:** Remote
**Role:** DevOps / SRE
**Salary:** $115k – $150k/yr
**Experience:** 5+ years
**Skills:** Site Reliability Engineering, Observability, Slis, SLOs, Error Budgets, Incident Management, Grafana, CloudWatch, SQL, Postgres, PHP, .Net, Node.js, Automation, Synthetic Monitoring
**Posted:** 2026-09-08

> Senior Site Reliability Engineer responsible for production troubleshooting, incident response, observability, SLOs, automation, and permanent reliability improvements. Requires strong software engineering, SQL, debugging, cloud-application troubleshooting, and cross-functional collaboration skills.

## Job Description

## Responsibilities
- Own complex production-support escalations and ticket triage, troubleshooting and resolving issues alongside reliability work.
- Partner with Software Engineering to investigate production issues, identify root causes and reliability risks, and drive permanent fixes.
- Apply and mature Site Reliability Engineering practices, including automation, continuous improvement, shared ownership, and toil reduction.
- Lead incident response through triage, mitigation, recovery, root-cause analysis, and blameless post-incident reviews.
- Build observability with meaningful metrics, logs, traces, dashboards, and actionable alerts.
- Define and mature service-level indicators (SLIs) and service-level objectives (SLOs), including error budgets.
- Develop synthetic monitoring for critical customer journeys.
- Automate recurring operational toil through tooling, process improvements, or permanent fixes.
- Use AI-assisted tools and source-code repositories for triage, troubleshooting, code analysis, automation, and investigation.
- Participate in a rotating on-call schedule, primarily during business hours, with limited after-hours and weekend support.

## Requirements
- Hands-on Site Reliability Engineering experience applying software engineering practices to production reliability and helping establish or mature SRE practices.
- Strong knowledge of SLIs, SLOs, error budgets, observability, automation, and toil reduction.
- Experience building monitoring, dashboards, alerts, and telemetry with tools such as Honeycomb, New Relic, Grafana, CloudWatch, Kibana, or similar.
- Experience with production incident management, root-cause analysis, blameless post-incident reviews, and corrective-action follow-through.
- Strong programming and scripting skills for troubleshooting application code and building automation and operational tooling.
- Strong SQL and relational-database skills for production troubleshooting and safe data correction; PostgreSQL preferred.
- Strong code literacy and debugging skills, including navigating unfamiliar codebases, understanding application flow, reviewing code and change history, and identifying reliability issues.
- Experience troubleshooting cloud-hosted applications using source code, logs, APIs, telemetry, event streams, and databases.
- Comfort navigating application stacks involving PHP, .NET, and Node.js; deep expertise in each is not required.
- Experience using AI-assisted tools in day-to-day engineering workflows.
- Strong communication and collaboration skills across Software Engineering, Product, Support, DevOps, and other technical teams.

## Compensation and Benefits
- Salary: $114,800–$150,000 annually.
- Eligibility for a discretionary bonus.
- Health, vision, and dental insurance, plus 24/7 healthcare access.
- 20 PTO days, 3 flex days, 4 volunteer days, 12 paid holidays, and paid parental leave.
- 401(k) match.
- Company-provided equipment.

## Similar jobs

- [Senior Network Engineer](https://hotfix.jobs/jobs/b23e626b-5598-4d12-964a-efca91f98bfd) - MongoDB - Palo Alto, CA - $118k – $231k/yr
- [Senior Site Infrastructure Engineer](https://hotfix.jobs/jobs/847213bd-91b8-44f9-b450-477e4aa3714e) - Shield AI - Seattle, WA - $110k – $210k/yr
- [Senior Site Reliability Engineer](https://hotfix.jobs/jobs/a6949459-b2bc-44d0-99bf-23c9890dcf26) - PrizePicks - Remote - $120k – $175k/yr
- [Lead Scientific Imaging Systems Engineer](https://hotfix.jobs/jobs/c09c30f6-dd1b-4de5-aeb3-248a726a5186) - Axle - Rockville, MD - $120k – $145k/yr
- [Senior DevOps Engineer](https://hotfix.jobs/jobs/861dce16-9869-482f-9c45-e406426a87f1) - Upstart - Remote - $136k – $197k/yr

**Apply:** https://hotfix.jobs/jobs/0cb9abcb-6159-44b6-a3ae-bef58f9fdbdb
**Canonical:** https://hotfix.jobs/jobs/0cb9abcb-6159-44b6-a3ae-bef58f9fdbdb