Skip to content
AnthropicAnthropic

Technical Program Manager, Compute

Drives planning, coordination, and execution of compute infrastructure programs at scale, owning workstreams from procurement to allocation. Partners cross-functionally with engineering, research, and finance teams in a fast-paced AI environment.

About the job

Responsibilities:

  • Own and drive critical programs across the compute lifecycle, coordinating execution across multiple engineering, research, and operations teams
  • Build and maintain operational visibility into the compute fleet, ensuring the organization has a clear picture of supply, demand, utilization, and health
  • Lead cross-functional coordination for compute transitions: bringing new capacity online, migrating workloads, and managing decommissions across cloud providers and hardware platforms
  • Partner with engineering and research leadership to navigate competing priorities and drive alignment on how compute resources are planned, allocated, and used
  • Identify and close operational gaps across the compute pipeline, whether through new tooling, improved processes, or better cross-team communication
  • Own trade-off discussions between utilization, cost, latency, and reliability, synthesizing inputs from technical and business stakeholders and communicating decisions to leadership
  • Develop and improve the processes and frameworks the team uses to plan, track, and execute compute programs at increasing scale and complexity

You may be a good fit if you:

  • Have 7+ years of technical program management experience in infrastructure, platform engineering, or compute-intensive environments
  • Have led complex, cross-functional programs involving multiple engineering teams with competing priorities and ambiguous requirements
  • Have experience working with research or ML teams and translating their needs into operational plans and technical requirements
  • Are comfortable diving deep into technical details (cloud infrastructure, cluster management, job scheduling, resource orchestration) while maintaining program-level visibility
  • Thrive in ambiguous, fast-moving environments where you need to define scope and build processes from the ground up
  • Have strong communication skills and can engage credibly with engineers, researchers, finance, and executive leadership
  • Have a track record of building trust with engineering teams and driving changes through influence rather than authority

Strong candidates may also have:

  • Experience managing compute capacity across multiple cloud providers (AWS, GCP, Azure) or hybrid cloud/on-premise environments
  • Familiarity with job scheduling, resource orchestration, or workload management systems (Kubernetes, Slurm, Borg, YARN, or custom schedulers)
  • Experience with GPU or accelerator infrastructure, including the unique challenges of large-scale ML training and inference workloads
  • Built or improved observability for infrastructure systems: dashboards, alerting, efficiency metrics, or cost attribution
  • Capacity planning experience including demand forecasting, cost modeling, or hardware lifecycle management
  • Scaled through hypergrowth in AI/ML, HPC, or large-scale cloud environments

Annual Salary: $365,000 — $435,000 USD

Skills

Kubernetes, AWS, GCP, Azure, Slurm, Borg, Yarn, GPU, Cluster Management, Job Scheduling

Anthropic

Anthropic

San Francisco, CA
Technical Program Manager, Enterprise Readiness
$365k+/yrHybrid8+ YOETechnical Program Management

Leads cross-functional enterprise-readiness programs spanning security, compliance, identity, data governance, and spend controls across product surfaces. Requires 8+ years of technical program management experience and strong experience with regulated customers and senior stakeholder alignment.

Anthropic

Anthropic

United States

Reporting and Controls Lead, Data Center Capacity Delivery
$320k+/yrHybrid7+ YOETechnical Program Management

Owns portfolio-level milestone tracking, reporting, data quality, and capacity projections for large-scale data center delivery programs. The role requires 7+ years in project controls or scheduling, strong scheduling and BI-tool fluency, and the ability to communicate insights to executive and technical audiences.

Anthropic

Anthropic

San Francisco, CA
Technical Program Manager, Billing
$290k+/yrHybrid6+ YOETechnical Program Management

Leads cross-functional programs for billing platform foundations, commercial launches, promotions, charge-pipeline changes, and payments operations. Requires 6+ years of technical program management experience and the ability to coordinate engineering, finance, treasury, product, and support teams.

Anthropic

Anthropic

San Francisco, CA
Incident Manager - Detection & Response
$290k+/yrHybrid7+ YOETechnical Program Management

Leads and scales the security incident management lifecycle for Detection & Response, serving as incident commander and driving post-incident accountability, trend analysis, systemic improvements, and cross-functional coordination. Requires 7+ years of relevant experience, strong analytical and communication skills, and a bachelor’s degree or equivalent experience.

Anthropic

Anthropic

Austin, TX
Technical Deployment Lead, Applied AI
$275k+/yrHybridTechnical Program Management

Leads end-to-end delivery of custom AI-agent deployments for enterprise customers in regulated industries, coordinating technical teams, executive stakeholders, product scoping, architecture decisions, and value measurement. Requires production AI/ML deployment experience, enterprise delivery expertise, and strong executive presence.