Senior Infrastructure Engineer
Senior individual contributor owning infrastructure stack areas, driving scope from ambiguity, setting technical direction via RFCs, leading on-call escalations, and mentoring engineers. Requires 5+ years production infra experience with expertise in Linux, Kubernetes, AWS, Terraform, and strong coding skills.
About the job
Responsibilities
- Develop and drive scope in one or more areas of the infrastructure stack. Take ambiguous problems from framing through design, implementation, and operation.
- Set the technical direction for your areas — write RFCs, evaluate tradeoffs, and make decisions that affect how the team operates for years.
- Act as the final escalation point on-call. Drive resolution of the hardest incidents and lead post-incident reviews that produce durable fixes.
- Raise the bar through design review, code review, and mentorship. Develop other engineers by working alongside them.
- Partner with security, compliance, product engineering, and ground station ops as a peer technical leader. Translate between infrastructure constraints and the needs of other teams.
- Identify systemic toil and underinvestment. Propose and execute the work that eliminates it, including work that spans other engineers' domains.
- Represent the infrastructure team in technical decisions across the organization. Write the documents others reference.
Basic Qualifications
- 5+ years of production infrastructure experience (SRE, DevOps, platform, network, or systems engineering), with demonstrated ownership of systems that were still operating well years after they were built.
- Working competence across: Linux, networking, AWS or equivalent cloud, infrastructure as code (Terraform), containers and Kubernetes, CI/CD pipelines, observability, secrets management.
- Expert-level depth in at least one of these.
- Track record of taking ambiguous problems and producing designs and implementation that others can operate.
- Senior on-call experience — has been the person a team escalates to.
- Strong code and automation skills (Python, Go, or similar).
- Proven technical writing and mentorship.
- US person status (ITAR requirement).
Preferred Qualifications
- Deep background in one or more of the following:
- Networking — enterprise (FortiGate, switching), site/ground station, or core (TGW, WireGuard, BGP, IPsec).
- Security engineering — Vault at scale, identity and IAM architecture, SIEM design and operation, hardening programs.
- Compliance partnership — leading technical controls implementation for CMMC, FedRAMP, NIST 800-171, ITAR.
- Observability at scale — designing SLO/SLI frameworks, operating VictoriaMetrics or Prometheus, building log pipelines.
- App platform — EKS at scale, on-prem k3s, multi-cluster ArgoCD and GitOps patterns.
- CI/CD and developer experience — building pipelines and platforms.
- Cloud architecture — multi-account AWS, GovCloud, Cloudflare, hybrid and air-gapped deployments.
- On-prem and hardware-adjacent ops.
Skills
Linux, Kubernetes, Terraform, AWS, Python, Go, CI/CD, Prometheus, EKS, Argo CD, Fortigate, Wireguard, BGP, Hashicorp Vault, Victoriametrics
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Senior software engineer building standardized, self-service cloud infrastructure across AWS, Google Cloud, and networking systems. Requires 5+ years of software engineering experience, production cloud infrastructure expertise, and proficiency in Go or Python.
Designs and supports physical IT infrastructure across offices, labs, manufacturing facilities, and data centers, including racks, cabling, power, cooling, documentation, and capacity planning. Requires 5+ years of physical infrastructure engineering experience and strong cross-functional project execution.