# Senior Platform Engineer

**Company:** [Shield AI](https://hotfix.jobs/companies/shield-ai)
**Location:** San Diego, CA, San Mateo, CA, Dallas, TX, Seattle, WA, Washington, DC, Boston, MA
**Role:** DevOps / SRE
**Salary:** $141k – $212k/yr
**Experience:** 7+ years
**Skills:** Azure, AWS, Terraform, Ansible, Kubernetes, Python, Go, Linux, CI/CD, Observability, Cloud Networking, Identity And Access Management, Helm, GitOps, VMware
**Posted:** 2026-09-09

> Designs and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.

## Job Description

## Responsibilities
- Design, build, and operate scalable platform infrastructure across Azure, AWS, and private cloud environments.
- Develop reusable infrastructure-as-code modules, automation frameworks, and self-service platform capabilities.
- Own platform initiatives from technical design through production operations and continuous improvement.
- Build and maintain container and Kubernetes platforms, including deployment automation, configuration management, observability, and lifecycle management.
- Develop platform tooling and automation using Terraform, Ansible, Python, and Go.
- Build standardized paved paths for infrastructure provisioning, application deployment, and shared platform services.
- Establish platform standards for reliability, scalability, security, performance, and cost efficiency.
- Support capacity planning, performance tuning, vulnerability remediation, disaster recovery, and platform lifecycle management.
- Build and improve CI/CD pipelines, monitoring, logging, alerting, and developer-facing platform services.
- Troubleshoot complex platform, infrastructure, and application integration issues and lead root-cause analysis.
- Maintain architecture diagrams, technical documentation, operational procedures, and reusable implementation patterns.
- Evaluate technologies and recommend improvements to automation, reliability, and developer productivity.
- Provide technical guidance and contribute to platform architecture, standards, and roadmap decisions.
- Participate in an on-call rotation and scheduled after-hours maintenance.

## Requirements
- 7+ years of experience in platform engineering, cloud infrastructure, DevOps, Site Reliability Engineering, or a related discipline.
- Experience designing and operating production infrastructure in Azure, AWS, or comparable public cloud environments.
- Strong experience with reusable infrastructure-as-code modules and automation using Terraform and Ansible.
- Hands-on experience deploying and operating containerized workloads and Kubernetes platforms.
- Proficiency in Python, Go, or another programming language used for platform tooling and automation.
- Experience building CI/CD pipelines, self-service infrastructure, or developer-facing platform services.
- Strong Linux systems administration experience, including deployment, configuration, troubleshooting, patching, and performance analysis.
- Knowledge of cloud and enterprise networking, including VPCs/VNets, subnets, routing, VPNs, load balancing, DNS, and firewalls.
- Experience implementing monitoring, logging, alerting, and observability for production platforms.
- Understanding of platform security, identity and access management, vulnerability remediation, and secure configuration practices.
- Ability to lead complex technical initiatives from design through production deployment.
- Strong technical documentation, communication, collaboration, and organizational skills.
- Bachelor's degree in computer science or a related field, or equivalent practical experience.

## Nice-to-haves
- Experience supporting commercial, government, regulated, classified, or air-gapped cloud environments.
- Experience with Microsoft Azure and AWS, including Azure Government or AWS GovCloud.
- Experience with GitOps, Helm, policy as code, secrets management, service catalogs, or internal developer portals.
- Experience designing or supporting internal developer platforms and standardized paved-path tooling.
- Knowledge of private cloud and virtualization platforms such as VMware, Hyper-V, or KVM.
- Experience defining service-level indicators, service-level objectives, incident response practices, and reliability improvements.
- Experience supporting hybrid-cloud architectures across public cloud, private cloud, and on-premises infrastructure.
- Experience in aerospace, defense, manufacturing, or other regulated environments.
- Relevant cloud, Kubernetes, Terraform, or platform-engineering certifications.

## Similar jobs

- [Senior Network Engineer](https://hotfix.jobs/jobs/d4ecdaa5-c0ab-49f3-baed-0b13deaaf6c0) - Shield AI - San Mateo, CA - $140k – $211k/yr
- [Senior Linux Infrastructure Engineer](https://hotfix.jobs/jobs/8d7d67b2-9c3a-4c7b-9f4b-4dd78b6fac23) - tastytrade - Chicago, IL - $140k – $180k/yr
- [Senior DevOps Engineer](https://hotfix.jobs/jobs/861dce16-9869-482f-9c45-e406426a87f1) - Upstart - Remote - $136k – $197k/yr
- [Senior Site Reliability Engineer](https://hotfix.jobs/jobs/b5f3d205-95e6-4655-8a41-935cc01b0e48) - Okta - San Francisco, CA - $147k – $227k/yr
- [Senior Site Reliability Engineer](https://hotfix.jobs/jobs/0a9de6d2-d738-400d-9f4e-a3fc9057c3e4) - Okta - Bellevue, WA - $147k – $202k/yr

**Apply:** https://hotfix.jobs/jobs/71245322-afd8-4019-8525-65088fed1493
**Canonical:** https://hotfix.jobs/jobs/71245322-afd8-4019-8525-65088fed1493