Software Engineer - Cloud Infrastructure
Builds and maintains large-scale cloud infrastructure for simulation workloads across AWS, GCP, and Azure. Requires 3+ years experience with Kubernetes, Docker, Golang/Python/C++, and container orchestration for high-reliability deployments.
About the job
Responsibilities
- Create and implement best practices for deploying and maintaining software for multiple customers across all three major cloud providers (AWS, GCP, Azure) with minimal downtime and high reliability
- Monitor Applied Intuition software on deployed machines and architect solutions to any bottlenecks that are encountered
- Improve developer efficiency by building internal tooling and optimizing our CI/CD systems
Requirements
- A Bachelor's degree in Computer Science, Software Engineering, or equivalent
- 3+ years of experience in a software engineering role working on infrastructure, monitoring, or a large scale software product
- 3+ years of coding experience from shell scripting (e.g., Bash) to higher-level languages (e.g., Python, C++)
- Experience with microservice orchestration frameworks such as Kubernetes
- Experience in the following: Golang, Python, Java, React, and C++
- Experience working with containerized systems (e.g., Docker)
Nice to Have
- Experience utilizing open source tooling
- Experience with deploying software on either public clouds (e.g., AWS) or on-premise clusters
Compensation
Base salary range: $126,000 - $186,000 USD annually (plus equity and benefits)
Skills
Kubernetes, Docker, Go, Python, Java, C++, React, AWS, GCP, Azure
Similar jobs
DevOps / SRE jobsOperate and scale Kong’s multi-region SaaS platform across major cloud providers, Kubernetes, and distributed data systems. The role requires strong infrastructure automation, observability, CI/CD, and production reliability experience, with participation in a global on-call rotation.
Builds and scales highly available infrastructure using AWS, Terraform, and Docker to support rapid growth and AI workloads. Collaborates with product and research teams on architectures, CI/CD, monitoring, and performance optimization.
Build and operate Mercor’s enterprise agent platform across security, routing, isolated execution, orchestration, deployment, and production scalability. The role requires 5+ years building high-scale platforms, architectural ownership, and experience with core infrastructure primitives across multiple clouds.
Owns secure, scalable Azure infrastructure for healthcare applications, including cloud migrations, Terraform-based automation, CI/CD pipelines, monitoring, and compliance. Requires 3–5+ years of Azure experience and strong DevOps and cloud-security expertise.
Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.