Skip to content
ReltioReltio

Staff Engineer, Release Management

Leads release management and cloud infrastructure initiatives for a highly available SaaS platform, mentoring the DevOps/Release team and improving automation, deployment, observability, security, and reliability. Requires 8+ years of enterprise SaaS development or operations experience and 6+ years with highly available cloud applications.

About the job

Responsibilities

  • Drive the planning, design, and build of effective solutions to support a rapidly growing business.
  • Lead initiatives supporting infrastructure for highly scalable and reliable platform services and components that form the foundation of the core platform architecture.
  • Mentor the DevOps/Release team, providing best practices and guidance to improve platform component performance.
  • Implement standards and best practices for the DevOps/Release team to maintain a 99.95% uptime SLA.
  • Collaborate with architects, engineers, product managers, and other partners to meet organizational needs from infrastructure through application layers.
  • Ensure compliance with security and privacy standards to support company certifications.
  • Continuously optimize and streamline infrastructure to provide a seamless customer experience.

Requirements

  • 8+ years of experience in enterprise SaaS software development and/or operations.
  • 6+ years of experience designing, deploying, and maintaining highly available cloud applications.
  • Ability to thrive in a dynamic, fast-paced environment, manage multiple responsibilities, prioritize effectively, and deliver results.
  • Experience with continuous testing, systems automation, virtualization, orchestration, continuous integration, deployment, and observability.
  • Knowledge of DevOps design patterns, processes, and best practices.
  • Understanding of build and deployment systems and system configuration.
  • Experience establishing end-to-end integrated delivery systems, including source control management, build tools, artifact repositories, deployment, and monitoring systems.
  • Strong verbal, written, and presentation skills.
  • Strong cross-team and cross-department partnership, collaboration, and consulting skills.
  • Strong focus on business outcomes.

Nice to Have

  • Understanding of data preparation, model training, and deployment processes.
  • Experience using AI/ML services from cloud providers, such as AWS SageMaker and Google AI Platform.

Skills

DevOps, Cloud Computing, Continuous Testing, Systems Automation, Virtualization, Orchestration, Continuous Integration, Continuous Deployment, Observability, Source Control Management, Build Tools, Artifact Repositories, System Configuration, Aws Sagemaker, Google Ai Platform

Datadog

Datadog

Dublin, Ireland
Staff Engineer, Compute
No salary listedHybrid7+ YOEDevOps / SRE

Leads the technical direction of multi-cloud Kubernetes capacity management and workload placement across Datadog’s large-scale infrastructure. The role requires strong systems programming experience, ideally in Go, cloud infrastructure expertise, and the ability to influence architecture across teams.

Fal

Fal

Remote

Senior/Staff Kubernetes Infrastructure Engineer
$180k+/yrRemote5+ YOEDevOps / SRE

Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.

Phantom

Phantom

Remote

Staff DevOps Engineer
No salary listedRemote8+ YOEDevOps / SRE

Owns and evolves CI/CD, mobile release, testing, and deployment infrastructure for a production fintech application. The role requires 8+ years in DevOps or related platform disciplines, strong AWS and Kubernetes expertise, and experience with secure mobile release systems.

Lightning AI

Lightning AI

Remote

Senior Network Engineer
$150k+/yrRemote5+ YOEDevOps / SRE

The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.

Cloudflare

Cloudflare

Lisbon, Portugal

Senior System Engineer
€66k+/yrHybrid6+ YOEDevOps / SRE

Build scalable infrastructure, automation, and network-resilience tools for a global cloud network. The role requires 6+ years in SRE, DevOps, or software engineering, strong Linux and distributed-systems expertise, programming skills in Python or Go, and experience with observability and IaC.