Director of Engineering, Compute Cloud
Leads the cloud compute engineering organization responsible for scalable IaaS, virtualization, and high-performance GPU infrastructure. The role requires extensive software engineering and management experience, expertise in distributed systems and compute technologies, and strong hiring and organizational leadership.
About the job
Responsibilities
- Lead cloud compute engineering at Crusoe Cloud and define the team’s long-term roadmap and strategy.
- Guide architecture decisions, design processes, design reviews, code reviews, and implementation.
- Scale the engineering organization through onboarding, hiring, manager development, and team processes.
- Track cloud software trends and incorporate relevant techniques into cloud offerings.
- Lead hiring and retention strategies for high-performing engineering talent.
Requirements
- 7+ years of engineering management experience and 10+ years of software engineering experience.
- Deep understanding of compute products, including CPU, GPU, and ASIC accelerators, and how customers consume them.
- Understanding of hypervisors and virtualization technologies such as KVM, VMware ESXi, and Microsoft Hyper-V.
- Familiarity with AI/ML workloads and their cloud data center compute requirements.
- Experience building and maintaining scalable, highly available, fault-tolerant distributed systems and application architectures.
- Knowledge of software engineering practices across the full software development life cycle, including coding standards, code reviews, source control, build processes, testing, and operations.
- Strong analytical, problem-solving, communication, and collaboration skills.
Nice-to-haves
- Experience operating production infrastructure against strict SLOs and SLAs.
- Experience with observability, monitoring, reliability, optimization, and capacity planning across large fleets.
- Expertise in Linux, configuration management, and deployment automation.
- Leadership in ambiguous, fast-moving environments with cross-functional teams.
- Open-source or upstream contributions to cloud infrastructure, containerization, or virtualization technologies such as KVM, QEMU, containerd, or Kubernetes.
Compensation and benefits
- $285,000–$335,000 annual compensation plus bonus.
- Restricted Stock Units included in all offers.
- Paid time off, paid holidays, and leave programs.
- Health, dental, and vision insurance; employer HSA contributions.
- Paid parental leave, life insurance, and short- and long-term disability coverage.
- Professional development and tuition reimbursement.
- Mental health and wellness support.
- Commuter benefits, cell phone stipend, and 401(k) with company match up to 4% of salary.
- Volunteer time off, global travel insurance, emergency assistance, daily meals allowance, and location-specific programs.
Skills
Cloud Computing, GPU, Cpu, Asic Accelerators, Kvm, Vmware Esxi, Microsoft Hyper-V, Virtualization, Distributed Systems, Linux, Kubernetes, Qemu, Containerd, Observability
Similar jobs
Engineering Management jobsLeads a team building foundational security services for Crusoe’s GPU cloud and infrastructure fleet, spanning identity, cryptography, runtime protection, access, vulnerability management, and architectural isolation. Requires 8+ years leading hands-on software or security engineering teams at large-scale infrastructure platforms.
Leads the architecture, production launch, and team buildout for a safety-critical flexible-compute control system that manages AI infrastructure power in coordination with utilities. Requires extensive software engineering and team leadership experience, distributed orchestration expertise, and knowledge of data-center power systems and grid programs.
Leads a 40-person applications engineering organization across the US and India, owning people development, architecture, delivery, quality, and technical strategy. Requires 10+ years of software engineering experience and 5+ years leading engineering teams and managers.
Leads the technical direction, engineering quality, reliability, and hands-on architecture of a 30-person organization building payment, ledger, wallet, and settlement infrastructure. Requires 10+ years of production software experience, 4+ years leading engineers, and deep distributed-systems and regulated-finance expertise.
Leads architecture, development, and optimization of global Order-to-Cash revenue systems, integrating billing, ERP, sales, and financial platforms. The role requires extensive revenue systems experience, strong data engineering and API expertise, and the ability to lead ERP transformations and technical teams.