Engineering Manager, Agent Orchestration
Leads Agent Orchestration team building core execution layer for AI agents, coordinating model reasoning, tool use, and evaluation at scale. Requires 2+ years engineering management, strong IC technical depth in distributed systems, and cross-functional collaboration.
About the job
In this role, you will
- Build, lead, and develop a high performing team of engineers, including hiring, coaching, and performance management.
- Own the technical strategy and roadmap for Decagon’s orchestration engine, balancing speed of iteration with correctness and safety.
- Drive architecture for systems that coordinate complex reasoning and action flows at scale.
- Set reliability, testing, and observability standards across the orchestration stack, and build an operating cadence that prevents repeated incidents.
- Create frameworks and guardrails that enable fast, safe iteration on agent behavior, evaluation, and rollout.
- Partner with Product, Research, and Infrastructure teams to define requirements, navigate tradeoffs, and ship multi-quarter initiatives that move core company metrics.
Your background will look something like this
- Have 2+ years of engineering management experience leading high performing teams in fast-moving environments.
- Have strong technical depth and an IC foundation that enables you to guide architecture, debug complex failures, and make sound trade-offs.
- Have experience building distributed systems, execution engines, real-time platforms, or other high scale systems where correctness and reliability matter.
- Have a track record of delivering multi-quarter projects through ambiguity, creating clarity for your team and stakeholders.
- Care deeply about engineering craft and operational excellence, including testing strategy, observability, and incident learning.
- Communicate clearly and collaborate well across Product, Research, and Infrastructure teams.
Even better if you have
- Experience with agent frameworks, runtimes, orchestration logic, or tool use systems.
- Experience with evaluation, experimentation, or model quality measurement systems.
- Experience building guardrails for safety-critical or highly reliable systems.
Compensation
$280,000 - $430,000 + equity
Skills
Distributed Systems, Agent Frameworks, Orchestration, LLMs, Observability, Testing, Real-Time Platforms, Evaluation Systems, Execution Engines
Similar jobs
Engineering Management jobsLeads and grows a research engineering team developing and productionizing conversational AI models and agent systems. The role requires substantial machine-learning systems experience, foundation-model expertise, and people-management experience.
Leads and remains hands-on with a 5–6-person team building data platform, ML infrastructure, and underwriting systems that power capital offers. Requires 5+ years of software engineering experience, technical leadership or management experience, and expertise in modern data and cloud infrastructure.
Leads and grows infrastructure engineering teams responsible for highly reliable, large-scale distributed systems and core production platforms. The role requires experience operating mission-critical services, building platform or storage infrastructure across multiple clouds, and strong technical and people leadership.
Leads and develops senior backend and platform engineers while owning architecture, reliability, and technical standards for systems moving money globally. The role requires at least five years of leadership-level experience, current hands-on engineering, distributed systems expertise, and high-stakes payments or fintech experience.
Leads the Web Infrastructure team responsible for Notion’s web client architecture, performance, reliability, and shared design systems. The role manages senior engineers and managers, drives execution and planning, and partners across the organization on technical and organizational practices.