Senior Infrastructure Engineer, DevOps
Own and scale AWS infrastructure, CI/CD pipelines, GitOps automation, observability, and cloud security for high-throughput payment systems. The role requires 5+ years of DevOps or infrastructure experience plus strong Kubernetes, automation, and SRE expertise.
About the job
Responsibilities
- Design, build, and optimize highly available AWS infrastructure managed through infrastructure as code.
- Own, modernize, and scale CI/CD pipelines and build self-service developer platforms and standardized golden paths.
- Implement self-healing infrastructure, GitOps workflows, automated deployment scripts, and reproducible environment templates.
- Configure and maintain observability tooling, including telemetry and monitoring for proactive bottleneck detection.
- Participate in on-call rotations and lead blameless postmortems to improve system resilience.
- Embed cloud security, compliance controls, API gateways, and network isolation into deployment workflows as code.
- Partner with product and engineering teams on platform adoption and mentor engineers on SRE, DevOps, and cloud-native practices.
Requirements
- 5+ years of hands-on experience in DevOps, platform engineering, or infrastructure roles.
- Experience owning complex, large-scale distributed systems in mission-critical environments.
- Advanced Kubernetes, including Amazon EKS, Docker, and GitOps deployment workflows.
- Experience with Argo CD, Flux, and GitHub Actions.
- Proven ability to build, optimize, and secure automated CI/CD pipelines.
- Strong coding or scripting skills in Python, Go, or Bash.
- Hands-on experience implementing telemetry and monitoring stacks using Prometheus, Grafana, OpenTelemetry, or similar tools.
- Drive to eliminate manual work and automate repetitive operational tasks.
Nice to Have
- Experience with tokenization, payment processing, or security products.
- Bachelor's degree.
- Experience managing distributed data streaming platforms such as Kafka or Amazon MSK.
- Experience operating high-concurrency database platforms under heavy load.
- Experience integrating AI developer tools, agentic workflows, or automated testing into CI/CD pipelines.
- Understanding of regulated security frameworks such as PCI-DSS, SOC 2, or ISO 27001.
- Ability to thrive in a fast-paced startup environment.
Compensation
- Annual salary: $80,000–$85,000.
Skills
AWS, Kubernetes, Amazon Eks, Docker, GitOps, Argo Cd, Flux, GitHub Actions, CI/CD, Python, Go, Bash, Prometheus, Grafana, OpenTelemetry
Similar jobs
DevOps / SRE jobsThe Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Designs and operates scalable, highly available cloud infrastructure while leading efficiency initiatives across compute, storage, networking, and cost optimization. Requires 5+ years of distributed-systems software development experience and expertise with cloud platforms, infrastructure as code, and Kubernetes.
Build and optimize ClickHouse Cloud’s highly available, multi-cloud infrastructure, including automation, distributed systems, networking, security, and cost-efficiency tooling. Requires 5+ years of experience operating scalable systems and expertise in cloud platforms, infrastructure as code, and production engineering.
Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.