DevOps Engineer
Owns and evolves cloud infrastructure, CI/CD, observability, developer tooling, and FinOps governance for a distributed engineering organization. Requires 6+ years in DevOps, SRE, or software engineering, plus strong Kubernetes, cloud, security, and automation experience.
About the job
Responsibilities
- Own and evolve the DevOps and platform engineering function, including build and deployment pipelines, monitoring, infrastructure as code, and cost optimization.
- Design and maintain secure, scalable CI/CD pipelines that support rapid product iterations without sacrificing stability.
- Improve observability using Datadog, GCP, and custom dashboards, including FinOps-focused observability outcomes.
- Implement and maintain FinOps governance, including showback and chargeback models, to drive cost accountability across engineering teams.
- Enable developer productivity by creating internal tooling, reusable modules, and simplified infrastructure.
- Help teams define and measure SLOs and establish standards for system health and reliability.
- Define cost-effective architectural patterns for new services and features with Product and Engineering.
- Shape engineering culture, DevOps practices, and the long-term DevOps roadmap.
- Collaborate with product engineering, security, finance, and other teams to align technical spending with business and financial goals.
Requirements
- 6+ years of experience in DevOps, SRE, or software engineering roles.
- Proficiency in Python or Go, with strong software engineering fundamentals.
- Experience owning production infrastructure and platform tooling, debugging complex distributed systems, and diagnosing unexpected financial spikes.
- Experience designing and implementing secure, robust infrastructure, particularly in environments with stringent regulatory requirements.
- Deep experience with Kubernetes and tools such as ArgoCD, Helm, and Terraform.
- Familiarity with GCP or other major cloud providers, with a focus on cloud cost management, reserved instances or savings plans, and cost anomaly detection.
- Hands-on experience with FinOps tooling such as Kubecost, AWS Cost Explorer, GCP Cloud Billing, or DoIT Cloud Intelligence, or experience building custom cost-allocation and reporting tools.
- Strong experience building CI/CD pipelines and internal developer tooling from the ground up.
- Hands-on experience with observability tools such as Prometheus or Datadog.
- Participation in an on-call rotation.
Compensation and Benefits
- Compensation in cash and equity.
- Early exercise for all options, including pre-vested options.
- Remote-first culture and work from anywhere.
- Flexible paid time off and year-end break.
- Health, dental, and vision coverage for employees and dependents, specific to the US and Canada.
- 4% 401(k) or RRSP matching, specific to the US and Canada.
- MacBook Pro delivered to your door.
- One-time home-office setup stipend.
- Monthly meal stipend.
- Monthly social meet-up stipend.
- Annual health and wellness stipend.
- Annual learning stipend.
Skills
Python, Go, Kubernetes, Argo CD, Helm, Terraform, GCP, Datadog, Prometheus, CI/CD, Finops, Kubecost, Infrastructure As Code, SLOs, Distributed Systems
Similar jobs
DevOps / SRE jobsLeads Ireland-based Platform Developer Enablement and SRE teams, defining platform strategy, developer self-service, reliability objectives, and observability standards. Requires senior software, SRE, or platform engineering experience, management leadership, and expertise in cloud infrastructure, Kubernetes, Terraform, CI/CD, and distributed systems.
Senior DevOps Engineer responsible for building and operating Kubernetes-based infrastructure, AWS cloud systems, deployment workflows, and observability for reliable services at scale. Requires 5+ years of DevOps or platform engineering experience and strong production Kubernetes expertise.
The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Operates and evolves high-throughput MariaDB infrastructure, improving reliability, automation, security, observability, and disaster recovery. Requires 5+ years of production MariaDB/MySQL experience plus expertise in distributed databases, Kubernetes, infrastructure as code, and incident readiness.