Engineer, Supercomputing & Distributed Systems
KreaSan Francisco, CA
Builds and operates supercomputing infrastructure for AI research including 1000+ GPU Kubernetes clusters, distributed data pipelines processing petabytes, and fault-tolerant training systems. Requires strong distributed systems intuition and experience with Python, PyTorch, and large-scale infrastructure.
Salary not listedOn-siteDevOps / SRE