Senior Software Engineer, AI Infrastructure
Senior engineer building and operating large-scale HPC infrastructure for AI model training. Owns job scheduling, automation, and performance optimization across GPU clusters.
Senior engineer building and operating large-scale HPC infrastructure for AI model training. Owns job scheduling, automation, and performance optimization across GPU clusters.
Builds and maintains automated IT infrastructure pipelines using Python, Terraform, and cloud providers to support company operations. Requires 5+ years experience with focus on automation, strong coding, and collaboration skills.
Builds foundational platform architecture for AI research agents, including SDKs, APIs, execution frameworks, and benchmarking infrastructure to enable researchers to develop intelligent systems over scholarly literature. Requires strong Python skills, 8+ years experience, cloud infrastructure, and AI integration expertise.
Leads AI infrastructure including on-prem GPU clusters, hybrid cloud orchestration, storage, and resource allocation to support frontier AI research. Requires 12+ years experience with HPC leadership, GPU systems, Kubernetes, and distributed storage.
Senior engineer building and operating large-scale HPC infrastructure for AI model training. Owns job scheduling, automation, and performance optimization across GPU clusters.
Builds and maintains automated IT infrastructure pipelines using Python, Terraform, and cloud providers to support company operations. Requires 5+ years experience with focus on automation, strong coding, and collaboration skills.
Builds foundational platform architecture for AI research agents, including SDKs, APIs, execution frameworks, and benchmarking infrastructure to enable researchers to develop intelligent systems over scholarly literature. Requires strong Python skills, 8+ years experience, cloud infrastructure, and AI integration expertise.
Leads AI infrastructure including on-prem GPU clusters, hybrid cloud orchestration, storage, and resource allocation to support frontier AI research. Requires 12+ years experience with HPC leadership, GPU systems, Kubernetes, and distributed storage.