# Senior Software Engineer, Infrastructure & Systems

**Company:** [Astronomer](https://hotfix.jobs/companies/astronomer)
**Location:** New York, NY
**Role:** DevOps / SRE
**Salary:** $200k – $300k/yr
**Experience:** 5+ years
**Skills:** Kubernetes, custom operators, crds, Helm, REST APIs, gRPC, Go, TypeScript, Distributed Systems, Networking, Observability, distributed tracing, Airflow, pod security standards
**Posted:** 2026-08-13

> Designs and operates control-plane systems that provision, scale, secure, and observe infrastructure running Airflow across multi-tenant and private-cloud environments. Requires 5+ years in infrastructure or systems engineering, strong Kubernetes and API expertise, and proficiency in Go or TypeScript.

## Job Description

## Responsibilities
- Design and own the architecture of systems that stand up, scale, configure, and manage the infrastructure running Airflow.
- Lead infrastructure efforts end-to-end, including design, implementation, testing, documentation, and rollout.
- Design and evolve REST and gRPC APIs for infrastructure lifecycle management, including data modeling, versioning, and backward compatibility.
- Own networking and security posture, including authentication, authorization, secure communication, CVE remediation, image hardening, and Pod Security Standards compliance.
- Architect observability and traceability across service and network boundaries using metrics, logs, and distributed traces.
- Participate in the on-call rotation, diagnose and resolve incidents, contribute to post-mortems, and track remediation work.
- Mentor engineers and improve design, code review, testing, and operational practices.

## Requirements
- 5+ years of experience in infrastructure, platform, or systems engineering, with a history of owning production systems at scale.
- Deep experience with Kubernetes, including custom Operators, CRDs, and Helm-based deployments.
- Experience designing REST or gRPC APIs for customer-facing interfaces or programmatic integrations, including versioning and backward compatibility.
- Strong fundamentals in networking, distributed systems, reliability engineering, and security.
- Proficiency in Go and/or TypeScript, or the ability to learn new programming languages quickly.
- Practical experience designing or building observability and distributed tracing for multi-layer systems.
- Ability to lead technical design across teams and communicate trade-offs to engineers and stakeholders.
- Experience mentoring engineers and improving team-wide practices.

## Nice to Have
- Experience designing or operating multi-tenant control planes.
- Experience shipping software into air-gapped or highly regulated environments.
- Familiarity with Apache Airflow internals or other data orchestration platforms.
- Incident command or on-call leadership experience.

## Compensation and Benefits
- Estimated total compensation: **$200,000–$300,000**, based on leveling and geography.
- Equity component and comprehensive benefits package.

## Similar roles

- [Senior Software Engineer - Snowpark Container Service](https://hotfix.jobs/jobs/47626f55-409c-4469-9cea-906cfd5683e3) - Snowflake - Bellevue, WA - $200k – $288k/yr
- [Lead Site Reliability Engineer](https://hotfix.jobs/jobs/00c74959-6a9f-4850-bc3a-75d81bdb364f) - Glean - Palo Alto, CA - $200k – $260k/yr
- [Senior Software Engineer, Snowpark Container Service](https://hotfix.jobs/jobs/3e6d8842-9131-4a2f-b8ef-77ac0972cbc6) - Snowflake - Bellevue, WA - $200k – $288k/yr
- [Senior Software Engineer, Cloud Infrastructure](https://hotfix.jobs/jobs/cd627f96-fe85-4e5c-988e-5705591da638) - Decagon - San Francisco, CA - $200k – $400k/yr
- [Senior Software Engineer, Observability](https://hotfix.jobs/jobs/b857fbeb-7984-45df-bd84-c85c53858f73) - Together AI - San Francisco, CA - $200k – $280k/yr

**Apply:** https://hotfix.jobs/jobs/23005e89-18ee-4619-858c-3e32bea46510
**Canonical:** https://hotfix.jobs/jobs/23005e89-18ee-4619-858c-3e32bea46510