# Senior Engineer, Platform Infrastructure

**Company:** [Shield AI](https://hotfix.jobs/companies/shield-ai)
**Location:** San Diego, CA, Washington, DC, California
**Role:** DevOps / SRE
**Salary:** $120k – $180k/yr
**Experience:** 5+ years
**Skills:** Kubernetes, Linux, Terraform, Ansible, Helm, zarf, packer, gitlab ci, Git, Containers, pki, secrets management, Observability, Distributed Systems, Infrastructure As Code
**Posted:** 2026-08-04

> Build and evolve the infrastructure platform that deploys and operates customer environments. The role focuses on Kubernetes, infrastructure as code, deployment automation, observability, reliability, security, and collaborative continuous delivery practices.

## Job Description

## How We Work

- Practice Extreme Programming and Continuous Delivery.
- Use pair programming as the default development approach.
- Work in small batches and integrate continuously.
- Automate repetitive work and test first whenever practical.
- Optimize for learning, feedback, and team outcomes.

## Responsibilities

- Build and improve the platform that deploys and operates customer environments.
- Develop infrastructure as code using Ansible, Terraform, Helm, Zarf, Big Bang, Packer, and related tooling.
- Improve Kubernetes platforms and surrounding systems.
- Build deployment automation that reduces risk and removes manual work.
- Pair with engineers to design, implement, and troubleshoot platform capabilities.
- Write automated tests for infrastructure and deployment workflows.
- Improve observability, reliability, security, and recoverability.
- Write documentation that explains why systems and processes exist.
- Decompose large problems into safe, incremental improvements.
- Investigate incidents beyond immediate fixes to prevent recurrence.

## Requirements

- Comfortable working across most of the following areas:
  - Kubernetes
  - Linux
  - Networking fundamentals
  - Git and trunk-based development
  - GitLab CI
  - Infrastructure as code
  - Ansible
  - Terraform
  - Helm
  - Zarf
  - Containers
  - PKI and certificate management
  - Secrets management
  - Observability
  - Infrastructure testing
  - Troubleshooting distributed systems
- Ability to learn quickly and become productive in unfamiliar systems.
- Collaborative approach to design, implementation, troubleshooting, and communicating tradeoffs.

## Success Measures

- Deploy more frequently and safely.
- Recover from failures faster.
- Remove manual work and reduce operational complexity.
- Improve documentation and confidence through testing.
- Deliver changes in small, reversible increments.
- Make the platform easier to operate and evolve.

## Values

- Simplicity over cleverness.
- Evidence over opinion.
- Learning over ego.
- Automation over repetition.
- Continuous improvement over perfection.
- Team outcomes over individual heroics.

## Similar roles

- [Senior Site Reliability Engineer](https://hotfix.jobs/jobs/a6949459-b2bc-44d0-99bf-23c9890dcf26) - PrizePicks - Remote - $120k – $175k/yr
- [Senior Network Engineer](https://hotfix.jobs/jobs/b80a174e-4942-4fa2-bd35-ad0184ece888) - CommandLink - Remote - $120k – $160k/yr
- [Senior Engineer, Software Engineering Tools (R4913)](https://hotfix.jobs/jobs/06ff6c1e-853f-4bd7-a741-7865d8b9c5c0) - Shield AI - Dallas, TX - $120k – $190k/yr
- [Senior Infrastructure Engineer](https://hotfix.jobs/jobs/c8f2b95b-8d32-4d8e-8ba2-0a29dfc5158a) - Bland AI - San Francisco, CA - $120k – $200k/yr
- [Senior Infrastructure Engineer](https://hotfix.jobs/jobs/d3847fd2-6f31-4a9c-bd48-e35523ad26a1) - LiveKit - Remote - $120k – $250k/yr

**Apply:** https://hotfix.jobs/jobs/cf15e8c3-ee2a-4f6a-abd6-8a54bce776fd
**Canonical:** https://hotfix.jobs/jobs/cf15e8c3-ee2a-4f6a-abd6-8a54bce776fd