Engineering Manager, Platform
Leads the platform engineering organization, managing managers and driving reliable, scalable infrastructure initiatives across the company. Requires 12+ years of engineering experience, substantial infrastructure leadership, AWS expertise, and experience with AI platforms, data infrastructure, and operational excellence.
About the job
Responsibilities
- Lead and scale a high-performing engineering team responsible for core platform systems used across the company.
- Define and drive multi-quarter platform initiatives that improve system reliability and product leverage.
- Partner with product and engineering leaders to evolve the platform as a product, balancing long-term architecture with near-term business needs.
- Turn ambiguous, high-impact problem areas such as system scalability, reliability, and platform abstractions into clear execution plans.
- Drive alignment across teams working on complex, interdependent systems, including internal platforms, infrastructure, and cloud governance.
- Raise the bar on engineering quality, operational excellence, and execution across the platform organization.
Requirements
- 12+ years of software or infrastructure engineering experience, including 5+ years leading infrastructure, SRE, or operations organizations and managing managers.
- Strong understanding of systems architecture.
- Bachelor's degree in Computer Science or equivalent experience.
- Ability to collaborate effectively cross-functionally with a strong sense of ownership.
- Deep experience with AWS at scale, including VPC, IAM, multi-account environments, core services, observability, and cost optimization.
- High AI fluency, including evaluating, integrating, and governing LLM and AI platforms such as AWS Bedrock, Anthropic/Claude, and OpenAI.
- Experience interfacing with vendors and partners and managing cost governance.
- Strong understanding of databases and data platforms, including OLTP, analytics, and data and ML/AI infrastructure.
- Proven record running high-reliability platforms, including SLOs, incidents, on-call processes, and postmortems.
- Experience driving processes and programs that improve engineering excellence.
- Clear communication skills and ability to manage escalations with CEO-level visibility.
- Track record of running infrastructure as a product, including roadmaps, stakeholder communications, and adoption.
Nice to Have
- Background in B2B SaaS or other high-reliability domains such as payments, logistics, or safety.
- Experience leading global, distributed teams across time zones.
Compensation and Benefits
- Base compensation range: $229,000–$290,000 USD.
- Benefits may include health, pharmacy, optical and dental care, paid time off, sick time off, short- and long-term disability coverage, life insurance, and 401(k) contributions, subject to eligibility requirements.
- Total compensation may include restricted stock units.
Skills
AWS, Amazon Vpc, Aws Iam, Cloud Governance, Observability, Llm Platforms, Aws Bedrock, Anthropic Claude, OpenAI, Oltp Databases, Data Platforms, Machine Learning Infrastructure, SLOs, Incident Management, Postmortems
Similar jobs
Engineering Management jobsLeads architecture and technical direction for a high-volume recognition platform spanning edge ingestion, real-time decisioning, identity, privacy, and operator systems. Requires 12+ years building distributed production systems and deep expertise in Scala or Java, streaming architectures, and cross-functional technical leadership.
Leads the engineering organization and technical strategy for Pinterest’s AI foundations, proactive experiences, and assistant capabilities. The role requires senior leadership experience managing managers, strong AI/platform depth, cross-functional execution, and a track record of building scalable, high-performing teams.
Staff Software Engineer leading technical strategy, complex AI-enabled systems, engineering mentorship, and privacy-focused practices. Requires 6+ years of software engineering experience, team leadership, and familiarity with AWS, Kubernetes, AI coding agents, data security, and HIPAA compliance.
Leads a small team responsible for the architecture, reliability, security, automation, and delivery of Anthropic’s global corporate campus and edge networks. The role requires 10+ years of enterprise networking experience, people management, and deep expertise in network operations and infrastructure.
Leads a mechanical engineering team developing and integrating flight- and mission-critical aircraft mechanisms and actuation systems for a military VTOL UAV. Requires a mechanical engineering degree, 10+ years of aerospace mechanisms experience, and demonstrated team leadership.