Engineering Manager - DevOps & AI Platform
Leads and develops DevOps and AI Platform teams responsible for reliable, scalable infrastructure and shared AI agent capabilities. The role requires people-management experience, strong cloud and Kubernetes knowledge, and practical curiosity about responsible AI adoption in engineering.
About the job
Responsibilities
- Lead, hire, level, coach, and develop a combined team of approximately 12 DevOps and AI Platform engineers.
- Set the DevOps strategy and operating model across release safety, developer experience, observability, cost management, and platform scalability.
- Lead the AI Platform centre of excellence, including shared agent-building frameworks and patterns, MCP management, evaluation frameworks, and shared AI infrastructure.
- Disseminate reusable AI and agent-building practices to product teams and channel feedback back into the centre of excellence.
- Partner with the AI Platform technical lead and senior engineers on technical direction, delivery, and execution.
- Own infrastructure cost, scalability, and efficiency, applying a FinOps mindset.
- Champion SRE and reliability practices across the business in partnership with the SRE function.
- Support infrastructure and process maturity for ISO 27001 and SOC 2 certifications in partnership with InfoSec.
- Establish AI SDLC practices and enable responsible, effective use of AI in engineering.
Requirements
- Experience managing DevOps, platform, or infrastructure engineers, including hiring and building growing teams.
- Strong infrastructure background spanning cloud, infrastructure as code, CI/CD, and observability.
- Demonstrated curiosity about and practical experience with AI, agents, and evaluation approaches.
- Interest in AI-assisted software development and the adoption, practices, and guardrails required to use it effectively.
- Hands-on Kubernetes experience and sufficient technical literacy to evaluate platform decisions.
- Ability to demonstrate practical, responsible AI use in an engineering context and establish team norms for it.
- Strong people leadership, strategic judgment, and ability to guide senior engineers without needing to be the hands-on technical expert.
Nice-to-haves
- Experience leading or building an internal platform, ML platform, developer experience, or enablement function.
- Exposure to agent frameworks, MCP, or evaluation frameworks.
- Experience working at a scaling company.
- Exposure to ISO 27001 and SOC 2 requirements.
- Experience with high-throughput, low-latency systems in regulated or financial environments.
- Experience in crypto or financial crime prevention.
Compensation and Benefits
- £500 remote-working budget.
- £1,000 learning and development budget.
- 25 days of annual leave plus bank holidays.
- Additional birthday leave.
- Enhanced parental leave, including 16 weeks of fully paid leave for eligible employees.
- Private health insurance.
- Mental health support.
- Life assurance at four times salary.
- Cycle to Work Scheme.
Skills
DevOps, Cloud Infrastructure, Infrastructure As Code, CI/CD, Observability, Kubernetes, SRE, Finops, AI Agents, Mcp, Evaluation Frameworks, Shared Ai Infrastructure, ISO 27001, SOC 2, Infrastructure Security
Similar jobs
Engineering Management jobsLeads and develops Dandy’s Sales Engineering team, partnering with Account Executives to support revenue growth through technical sales, product demonstrations, and customer advisory. Requires at least five years of dental-industry Sales Engineering experience and strong people leadership skills.
Leads and grows a globally distributed engineering team responsible for high-performance trading backend services. The role requires technical management experience, hands-on expertise with concurrent or low-latency systems, and proficiency in Rust or C++.
Leads GitLab’s Build team, owning the strategy, delivery, security, and operational excellence of systems that produce distributable software artifacts. The role requires distributed-systems expertise, experience operating highly available services, and a track record managing high-performing remote engineering teams.
Leads technical strategy and a multidisciplinary engineering team responsible for billing, accounting, eligibility, and enrollment systems. The role requires 5+ years of engineering management experience, strong distributed-systems expertise with Python or Go, and experience developing engineering leaders.
Leads two engineering squads building data infrastructure for global football and rugby analytics products. The role combines people management, platform and API architecture, product delivery, and AI-first engineering, requiring at least three years of engineering management experience and strong technical depth.