Skip to content
Cerebras SystemsCerebras SystemsSunnyvale, CA

Director/Sr. Manager, AI Inference Model Scaling

Leads a globally distributed engineering organization enabling foundation models and generative AI workloads on Cerebras hardware. The role requires deep compiler or ML infrastructure expertise, substantial engineering leadership experience, and cross-functional collaboration across hardware, runtime, cloud, and research teams.

Salary not listed
Hybrid12+ YOEEngineering Management

About the role

Responsibilities

Technical Leadership

  • Define the technical roadmap and strategy for the team.
  • Establish technical direction across multiple teams and engineering leaders.
  • Lead design reviews and establish engineering standards.
  • Drive support for emerging LLM architectures and inference workloads.

Team Leadership

  • Hire, mentor, and grow a high-performing engineering team.
  • Develop future technical leaders and managers.
  • Drive organizational planning, headcount strategy, and investment priorities.
  • Foster a strong engineering culture focused on execution, quality, and innovation.
  • Scale engineering processes while maintaining execution velocity.

Cross-Functional Collaboration

  • Partner with Cloud Platform, ML, and Hardware teams in planning and delivering end-to-end service enablement in Cloud and On-Premise settings.
  • Work with Product Management to prioritize model enablement and customer needs.
  • Collaborate closely with customers and solution architects on new model bring-up.
  • Influence future hardware/software co-design through ML model enablement and optimization insights.

Delivery and Execution

  • Own planning, prioritization, and execution across multiple concurrent initiatives.
  • Balance rapid model support with long-term ML Compiler architecture.
  • Drive predictable delivery for strategic customer commitments.

Required Qualifications

  • BS, MS, or PhD in Computer Science, Computer Engineering, or a related field.
  • 12+ years building compiler, ML systems, or infrastructure software.
  • 5+ years leading engineering teams.
  • Deep experience with modern compiler infrastructure, such as LLVM, MLIR, XLA, TVM, or Torch FX.
  • Strong understanding of graph compilation and optimization.
  • Experience with Python and C++.
  • Experience delivering production-quality software.
  • Strong communication and cross-functional leadership skills.

Preferred Qualifications

  • Experience building compiler frontends for AI accelerators.
  • Experience supporting PyTorch, JAX, TensorFlow, or ONNX.
  • Experience with LLM inference or training systems.
  • Familiarity with distributed compilation.
  • Experience working with hardware architects.
  • Experience leading teams through rapid growth.

Benefits and Culture

  • Build a breakthrough AI platform beyond the constraints of GPUs.
  • Publish and open source cutting-edge AI research.
  • Work on one of the fastest AI supercomputers in the world.
  • Enjoy job stability with startup vitality.
  • Work in a simple, non-corporate culture that respects individual beliefs.

Skills

C++Pythonllvmmlirxlatvmtorch fxPyTorchJAXTensorFlowonnxgraph compilationDistributed Systemsllm inference
Nourish

Senior Manager / Director of Engineering, Consumer

NourishNew York, NY +1

Leads Nourish’s consumer engineering organization, setting technical and product strategy while managing multiple teams, hiring and developing engineers, and remaining hands-on with architecture and code. Requires substantial software engineering and engineering management experience, strong product judgment, and proficiency with modern full-stack technologies.

Salary not listedOn-site8+ YOEEngineering Management
Snowflake

Director of Engineering, AI Enterprise

SnowflakeMenlo Park, CA +1

Leads a global AI Solutions engineering organization, setting technical vision and delivering enterprise AI/ML solutions on Snowflake’s Cortex platform. The role requires 12+ years of software engineering experience, director-level leadership managing managers, deep generative AI expertise, and strong enterprise customer and cross-functional partnership skills.

264k – 380k/yrHybrid12+ YOEEngineering Management
LeafLink

Head of Engineering

LeafLinkUnited States

Leads the entire engineering organization, owning technical vision, architecture, execution standards, security, budget, and organizational scaling. The role requires extensive software engineering and engineering management experience, cloud-native architecture expertise, and a record of building high-performing teams.

240k – 270k/yrRemote12+ YOEEngineering Management
Motive

Director, Developer Platform & Experience

MotiveSan Francisco, CA +1

Leads a multi-team organization responsible for AI-assisted development, cloud agentic infrastructure, developer experience, CI/CD, testing, and engineering velocity. Requires extensive software engineering and engineering leadership experience, deep infrastructure expertise, and hands-on knowledge of AI developer tooling and LLM evaluation.

229k – 285k/yrHybrid12+ YOEEngineering Management
OPSWAT

Director of Engineering

OPSWATUnited States

Hands-on Director of Engineering leading AI-native development for MetaDefender Email Security (on-prem gateway and cloud M365). Owns architecture, delivery, quality, and teams; orchestrates AI agents across full SDLC while maintaining human oversight and high security bar. Requires proven agentic product launches in high-assurance environments plus 8+ years engineering and leadership experience.

Salary not listedRemote8+ YOEEngineering Management