Infrastructure Software Engineer
Build and operate the cloud and edge infrastructure supporting a large IoT product fleet, including AWS resources, device management, CI/CD, observability, and automation. The role requires at least three years of infrastructure and programming experience, with AWS, Kubernetes, and IoT expertise.
About the job
Responsibilities
- Write and maintain production-grade software for the custom edge infrastructure stack.
- Provision and maintain resources running on AWS.
- Build provisioning and management systems for hundreds of thousands of connected IoT devices deployed in the field.
- Build CI/CD and automation pipelines for various parts of the stack.
- Build observability and telemetry across cloud applications and edge devices.
- Help maintain compliance with security standards, including SOC 2 and HIPAA.
- Maximize developer productivity by streamlining development workflows.
Requirements
- 3+ years of experience writing production infrastructure running on AWS using infrastructure-as-code tools such as Pulumi or Terraform.
- Experience with Docker and Kubernetes, particularly Amazon EKS.
- 3+ years of experience with Python, Go, or another modern programming language.
- Experience building CI/CD pipelines and automating parts of the stack.
- Experience self-hosting and maintaining observability tools such as Grafana and Prometheus.
- Experience with edge/IoT infrastructure, including Yocto, IoT device provisioning, and over-the-air updates.
- Experience remotely managing on-premises infrastructure or hybrid/multi-cloud environments.
- Experience maintaining SOC 2 compliance.
- Experience building high-volume data-processing pipelines.
- Resilience in challenging, fast-paced environments.
- Exceptional written and verbal communication skills in English, with the ability to influence at all levels.
- Ability to work onsite.
Benefits
- Flexible paid time off and paid holidays.
- Early-stage equity in a rapidly growing company.
- Referral bonuses.
- Regular team off-sites.
- Latest Apple products and access to a broad technology stack.
Skills
AWS, Pulumi, Terraform, Docker, Kubernetes, Amazon Eks, Python, Go, CI/CD, Grafana, Prometheus, Yocto, Iot, SOC 2, HIPAA
Similar jobs
DevOps / SRE jobsBuild and maintain Cloudflare’s deployment platform, enabling progressive rollouts, health-mediated releases, and automated workflows at scale. The role requires at least four years of software development experience, backend and frontend experience, and comfort with rapid delivery and on-call support.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Site Reliability Engineers build and operate reliable, scalable production infrastructure across GitLab’s Infrastructure Platforms teams. The role requires strong software engineering and operations fundamentals, Kubernetes and infrastructure-as-code experience, cloud expertise, and comfort with automation, observability, and incident response.
Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Infrastructure engineer responsible for building and operating highly available cloud systems, automating operations, and improving reliability across a large-scale AI platform. Requires 5+ years of infrastructure or DevOps experience, production Kubernetes, cloud infrastructure, Terraform, and Python or Go.