Skip to content
VGSVGS

Senior Infrastructure Engineer, DevOps

Own and scale AWS infrastructure, CI/CD pipelines, GitOps automation, observability, and cloud security for high-throughput payment systems. The role requires 5+ years of DevOps or infrastructure experience plus strong Kubernetes, automation, and SRE expertise.

About the job

Responsibilities

  • Design, build, and optimize highly available AWS infrastructure managed through infrastructure as code.
  • Own, modernize, and scale CI/CD pipelines and build self-service developer platforms and standardized golden paths.
  • Implement self-healing infrastructure, GitOps workflows, automated deployment scripts, and reproducible environment templates.
  • Configure and maintain observability tooling, including telemetry and monitoring for proactive bottleneck detection.
  • Participate in on-call rotations and lead blameless postmortems to improve system resilience.
  • Embed cloud security, compliance controls, API gateways, and network isolation into deployment workflows as code.
  • Partner with product and engineering teams on platform adoption and mentor engineers on SRE, DevOps, and cloud-native practices.

Requirements

  • 5+ years of hands-on experience in DevOps, platform engineering, or infrastructure roles.
  • Experience owning complex, large-scale distributed systems in mission-critical environments.
  • Advanced Kubernetes, including Amazon EKS, Docker, and GitOps deployment workflows.
  • Experience with Argo CD, Flux, and GitHub Actions.
  • Proven ability to build, optimize, and secure automated CI/CD pipelines.
  • Strong coding or scripting skills in Python, Go, or Bash.
  • Hands-on experience implementing telemetry and monitoring stacks using Prometheus, Grafana, OpenTelemetry, or similar tools.
  • Drive to eliminate manual work and automate repetitive operational tasks.

Nice to Have

  • Experience with tokenization, payment processing, or security products.
  • Bachelor's degree.
  • Experience managing distributed data streaming platforms such as Kafka or Amazon MSK.
  • Experience operating high-concurrency database platforms under heavy load.
  • Experience integrating AI developer tools, agentic workflows, or automated testing into CI/CD pipelines.
  • Understanding of regulated security frameworks such as PCI-DSS, SOC 2, or ISO 27001.
  • Ability to thrive in a fast-paced startup environment.

Compensation

  • Annual salary: $80,000–$85,000.

Skills

AWS, Kubernetes, Amazon Eks, Docker, GitOps, Argo Cd, Flux, GitHub Actions, CI/CD, Python, Go, Bash, Prometheus, Grafana, OpenTelemetry

Lightning AI

Lightning AI

Remote

Senior Network Engineer
$150k+/yrRemote5+ YOEDevOps / SRE

The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.

Lightspark

Lightspark

Remote

Senior Production Engineer
$200k+/yrRemote5+ YOEDevOps / SRE

The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.

Clickhouse

Clickhouse

Remote

Senior Cloud Software Engineer - Efficiency Engineering
No salary listedRemote5+ YOEDevOps / SRE

Designs and operates scalable, highly available cloud infrastructure while leading efficiency initiatives across compute, storage, networking, and cost optimization. Requires 5+ years of distributed-systems software development experience and expertise with cloud platforms, infrastructure as code, and Kubernetes.

Clickhouse

Clickhouse

Remote

Senior Cloud Software Engineer - Efficiency Engineering
No salary listedRemote5+ YOEDevOps / SRE

Build and optimize ClickHouse Cloud’s highly available, multi-cloud infrastructure, including automation, distributed systems, networking, security, and cost-efficiency tooling. Requires 5+ years of experience operating scalable systems and expertise in cloud platforms, infrastructure as code, and production engineering.

Fal

Fal

Remote

Senior/Staff Kubernetes Infrastructure Engineer
$180k+/yrRemote5+ YOEDevOps / SRE

Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.