Senior Manager, Platform Engineering
Leads Ireland-based Platform Developer Enablement and SRE teams, defining platform strategy, developer self-service, reliability objectives, and observability standards. Requires senior software, SRE, or platform engineering experience, management leadership, and expertise in cloud infrastructure, Kubernetes, Terraform, CI/CD, and distributed systems.
About the job
Responsibilities
Leadership & Strategy
- Lead, inspire, and manage the Platform Developer Enablement and Site Reliability Engineering teams.
- Set goals, hire and develop talent, and mentor engineering managers.
- Define the platform roadmap and align technical investments with business objectives.
- Champion DevOps, automation, and infrastructure-as-code practices.
Platform Development & Operations
- Oversee the design and operation of the platform for product teams and internal stakeholders.
- Drive architectural evolution to maintain scalability, resilience, and security.
- Establish and maintain rigorous service-level objectives (SLOs).
- Implement observability, monitoring, and logging frameworks.
Collaboration & Communication
- Partner with product management, security, and engineering leaders to translate business needs into platform capabilities.
- Communicate technical risks and architectural trade-offs to technical and non-technical leadership.
- Promote platform design patterns that balance velocity, reliability, and security.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
- 5+ years of senior-level experience as a Software Engineer, SRE, and/or Platform Engineer.
- 5+ years of management experience, including leading and coaching managers.
- Experience running SaaS products in cloud environments; AWS preferred.
- Strong experience with Kubernetes, Terraform, and modern CI/CD pipelines, including GitOps and ArgoCD.
- Programming experience in Go, Java, or a similar language for technical guidance and code reviews.
- Expertise in distributed systems design and architecture patterns.
- Advanced observability knowledge, particularly OpenTelemetry.
- Familiarity with MySQL, PostgreSQL, DynamoDB, MongoDB, Kafka, and Redis.
- Excellent communication skills and the ability to connect technical execution with business strategy.
Compensation & Benefits
- Anticipated salary range: EUR 120,000–140,000, plus variable commission or company bonus.
- Equity plan and bonus plan for eligible roles.
- Modern Health financial, mental, and physical wellness support.
- Retirement plan and match for US offices; local country pension for international offices.
- Unlimited vacation and flexible hours.
- Comprehensive medical benefits.
- Employee assistance and wellness reimbursement programs.
Skills
AWS, Kubernetes, Terraform, CI/CD, GitOps, Argo CD, Go, Java, Distributed Systems, OpenTelemetry, MySQL, Postgres, DynamoDB, MongoDB, Kafka
Similar jobs
DevOps / SRE jobsSenior DevOps Engineer responsible for building and operating Kubernetes-based infrastructure, AWS cloud systems, deployment workflows, and observability for reliable services at scale. Requires 5+ years of DevOps or platform engineering experience and strong production Kubernetes expertise.
The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Operates and evolves high-throughput MariaDB infrastructure, improving reliability, automation, security, observability, and disaster recovery. Requires 5+ years of production MariaDB/MySQL experience plus expertise in distributed databases, Kubernetes, infrastructure as code, and incident readiness.
Operates and scales Crusoe Cloud’s global edge, backbone, and data center networks supporting GPU-based HPC workloads. The role requires extensive production networking experience, strong protocol and observability expertise, automation skills, and participation in 24/7 on-call support.