Owns and evolves CI/CD, mobile release, testing, and deployment infrastructure for a production fintech application. The role requires 8+ years in DevOps or related platform disciplines, strong AWS and Kubernetes expertise, and experience with secure mobile release systems.
Salary not listed
Remote8+ YOEDevOps / SRE
About the role
Responsibilities
Own and evolve SDLC infrastructure using GitHub Actions and Kubernetes, focusing on reliability, scalability, and developer velocity.
Automate mobile release workflows, including building, signing, distributing binaries, and managing app store submissions across iOS, Android, and Chrome.
Oversee and improve automated end-to-end testing with Maestro and Playwright across mobile and web.
Improve pipeline performance, reliability, and observability to support rapid, high-confidence releases.
Partner with engineering teams on SDLC and developer experience improvements, including preview environments and Appetize tooling.
Help ensure build, release, and deployment systems meet the security and compliance expectations of a financial product.
Set standards and drive improvements across build, test, release, and deployment workflows in partnership with product and engineering teams.
Requirements
8+ years of experience in DevOps, Platform Engineering, SRE, Systems Engineering, or a related role, including experience operating at a senior or staff level.
Strong experience with AWS and CI/CD systems such as GitHub Actions or GitLab CI.
Strong experience with Docker, Kubernetes, and infrastructure-as-code tools such as Terraform or Pulumi.
Experience supporting CI/CD and release infrastructure for production mobile applications, including iOS and Android.
Experience working in financial services, fintech, or another environment with meaningful security and compliance requirements.
Strong track record of improving developer workflows and partnering across engineering teams.
Nice to Have
Experience with React Native and/or Expo, mobile preview environments, or Appetize.
Blockchain or crypto experience.
Compensation and Benefits
Competitive salary and equity.
Medical, dental, and vision insurance fully covered.
Stipend for remote-work equipment and setup.
Flexible hours and a supportive remote environment.
Build and operate Reddit’s internet-scale observability platform across monitoring, logging, and distributed tracing. The role requires 7+ years of infrastructure or software engineering experience, distributed systems expertise, and strong Kubernetes and troubleshooting skills.
217k – 304k/yrRemote7+ YOEDevOps / SRE
Senior/Staff Kubernetes Infrastructure Engineer
FalUnited States
Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.
180k – 250k/yrRemote5+ YOEDevOps / SRE
Senior Staff Deployment Automation Engineer
CrusoeSan Francisco, CA +2
Owns deployment, CI/CD, and integration-testing automation for large-scale multi-node GPU and CPU clusters. The role requires 12+ years of experience, strong Python or Bash skills, and expertise across Linux, Kubernetes, configuration management, GPU ecosystems, and high-performance networking.
250k – 300k/yrOn-site12+ YOEDevOps / SRE
Staff Site Reliability Engineer
SkydioUnited States
Owns and scales production cloud infrastructure across Kubernetes/EKS, AWS, Terraform, CI/CD, networking, and observability. The role requires 8+ years of infrastructure experience, strong Kubernetes operations expertise, and depth in reliability or scaling challenges.
240k – 300k/yrRemote8+ YOEDevOps / SRE
Staff Site Reliability Engineer
AttentiveUnited States
Leads strategic production engineering initiatives that improve the reliability, scalability, observability, and security of large-scale platforms. The role requires 7+ years of relevant experience, strong coding skills, and expertise in reliability practices such as SLIs, SLOs, and incident management.