Skip to content
Regal.aiRegal.ai

Senior DevOps Engineer

Senior DevOps Engineer designs and operates AWS-based internal platforms for deployment, observability, and AI-assisted workflows. Requires 4+ years experience with Terraform, containers, serverless, CI/CD, and distributed systems reliability.

About the job

Responsibilities

  • Design and maintain Infrastructure as Code using Terraform to provision and manage AWS infrastructure.
  • Build and evolve internal developer platforms (IDPs) with self-service workflows for deployment, environment provisioning, and operational tooling.
  • Support self-serve developer environments for engineers and AI agents.
  • Enable and scale AI-assisted engineering operations, including access management and developer productivity measurement.
  • Improve system reliability, scalability, and performance across distributed services and real-time workloads.
  • Optimize CI/CD systems (GitHub Actions) to reduce deployment friction.
  • Operate production systems with ownership of uptime, incident response, and recovery automation.
  • Partner with engineering teams to improve service architecture, resiliency, and operational maturity.
  • Improve cost efficiency and infrastructure utilization.
  • Support secure and compliant cloud infrastructure (SOC2).

Requirements

  • 4+ years of experience in DevOps, Platform Engineering, or Cloud Infrastructure roles.
  • Deep experience operating production workloads on AWS.
  • Strong hands-on experience with Terraform or comparable IaC tooling.
  • Experience running containerized (ECS Fargate) and serverless (Lambda) systems at scale.
  • Strong understanding of distributed systems reliability and failure modes.
  • Experience building CI/CD pipelines and developer automation tooling.
  • Experience with observability platforms such as Datadog, Prometheus, or OpenTelemetry.
  • Familiarity with incident response, on-call operations, and production debugging.
  • Working knowledge of networking, IAM, and cloud security best practices.

Nice to Haves

  • Experience building or operating Internal Developer Platforms (IDPs).
  • Experience with self-serve developer environments at scale.
  • Exposure to AI coding tools and infrastructure.
  • Experience supporting AI/ML or real-time workloads.
  • Cost optimization and cloud efficiency initiatives.
  • Exposure to compliance frameworks (SOC2, HIPAA, PCI).

Benefits

  • Medical, Dental, and Vision plans - 80% covered.
  • Flexible PTO & 11 paid holidays.
  • 401k Plan.
  • Paid parental leave.
  • Pre-tax commuter benefits.
  • In-office breakfast and snacks, happy hours, team outings.

Skills

AWS, Terraform, Python, TypeScript, Ecs Fargate, AWS Lambda, Datadog, GitHub Actions, Aurora Postgres, DynamoDB, Kinesis, SQS, Kubernetes, SageMaker

Shield AI

Shield AI

San Mateo, CA
Senior Network Engineer
$140k+/yrOn-site6+ YOEDevOps / SRE

Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.

tastytrade

tastytrade

Chicago, IL

Senior Linux Infrastructure Engineer
$140k+/yrHybrid6+ YOEDevOps / SRE

Own and improve the Linux production infrastructure layer, from performance tuning and incident response to configuration management, orchestration, networking, virtualization, secrets, and observability. The role requires 6+ years of infrastructure or SRE experience and deep Linux expertise.

Shield AI

Shield AI

San Diego, CA
Senior Platform Engineer
$141k+/yrHybrid7+ YOEDevOps / SRE

Designs and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.

Upstart

Upstart

United States

Senior DevOps Engineer
$136k+/yrRemote3+ YOEDevOps / SRE

Build and operate developer platform systems for continuous integration, Kubernetes-based ephemeral environments, automated testing, and internal tooling. The role requires a bachelor’s degree or equivalent, three years of software engineering experience, and experience operating production software or infrastructure.

Okta

Okta

San Francisco, CA

Senior Site Reliability Engineer
$147k+/yrHybrid5+ YOEDevOps / SRE

Senior Site Reliability Engineer responsible for operating and improving large-scale, FedRAMP-compliant cloud services through automation, observability, incident response, and platform engineering. Requires strong Kubernetes, cloud infrastructure, software engineering, and reliability engineering expertise.