Software Engineer, Infrastructure
Builds scalable infrastructure platforms for data collection systems, including orchestration, dev platforms, and multi-tenant data lakes supporting exabyte-scale robotics data. Requires Python fluency and experience with distributed systems; AI/robotics background preferred.
About the job
Responsibilities
- Build platform supporting ever-increasing scale of data collection.
- Build orchestration system to run computations upon data ingestion.
- Design internal dev platform for researchers to test ideas over the dataset.
- Manage multi-tenant data lake.
Requirements
- Bachelor's degree or equivalent experience in Computer Science or related field.
- Fluency with Python.
Nice-to-Haves
- Previous experience with large-scale data processing and distributed systems.
- Prior experience with AI or robotics.
- Comfortable working in 0->1 environments.
- Mission-driven and passionate about robotics.
Skills
Python, Distributed Systems, Data Processing, Data Lake, Orchestration, Dev Platform, AI, Robotics
Similar jobs
DevOps / SRE jobsBuild and maintain Cloudflare’s deployment platform, enabling progressive rollouts, health-mediated releases, and automated workflows at scale. The role requires at least four years of software development experience, backend and frontend experience, and comfort with rapid delivery and on-call support.
Own large-scale ClickHouse cluster upgrades and production operations while building tooling that improves release safety and automation. The role requires 5+ years operating stateful distributed systems, cloud and Kubernetes experience, strong debugging skills, and Go development experience.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.