# AI Infrastructure Engineer, Sandbox Platform

**Company:** [Scale AI](https://hotfix.jobs/companies/scale-ai)
**Location:** San Francisco, CA, Seattle, WA, New York, NY, London, United Kingdom
**Role:** Backend Engineering
**Experience:** 4+ years
**Skills:** Linux, Go, Rust, C/C++, Docker, Firecracker, Gvisor, Qemu, Kubernetes, Cgroups, Namespaces, Criu, Virtualization, API Design, Llm Agents
**Posted:** 2026-08-04

> Build and operate a secure, high-performance sandboxing platform for agentic code execution across containerized and virtualized environments. The role requires systems software experience, deep Linux knowledge, proficiency in Go, Rust, or C/C++, and strong developer-facing API and production debugging skills.

## Job Description

## Responsibilities
- Design and build the sandboxing platform, client library, and API surface for secure code execution across containerized and virtualized environments.
- Ensure strong isolation, security, and reproducibility across user sessions and workloads.
- Optimize cold-start latency, memory footprint, and resource utilization at scale.
- Reduce error rates through systematic debugging, monitoring, and proactive fixes.
- Partner with internal teams to understand platform needs, debug issues, and build supporting tooling.
- Respond to incidents and production issues, conduct root-cause analysis, and implement preventive fixes.
- Help develop and maintain the sandboxing product roadmap, balancing immediate needs with long-term architecture.
- Lead architecture reviews and own projects end-to-end from design through deployment.

## Requirements
- 4+ years of experience building high-performance systems software, including meaningful experience maintaining libraries, SDKs, or developer-facing APIs.
- Deep understanding of Linux internals, including process isolation, memory management, cgroups, and namespaces.
- Experience with containerization and virtualization technologies such as Docker, Firecracker, gVisor, QEMU, or Kata Containers.
- Proficiency in a systems programming language such as Go, Rust, or C/C++.
- Strong focus on developer experience, including API design, error propagation, documentation, and library quality.
- Ability to work across infrastructure layers, from kernel modules to orchestration frameworks such as Kubernetes.
- Strong debugging skills and ability to navigate performance and security tradeoffs in production systems.
- Comfort with ambiguity and switching between incident response and proactive product development.

## Nice to Have
- Experience as a founder or early infrastructure startup engineer with end-to-end product ownership.
- Familiarity with LLM agents and agent frameworks such as OpenHands, Agent2Agent, or MCP.
- Experience running secure workloads in multi-tenant or untrusted environments, including FaaS, CI sandboxes, or remote notebooks.
- Exposure to snapshotting and restore techniques such as CRIU, VM snapshots, or overlays.
- Open-source contributions to systems or developer-tools projects.
- Production on-call or incident-response experience.

## Similar jobs

- [Software Engineer III](https://hotfix.jobs/jobs/93d9d506-bc88-44b7-9876-42e5eba635c9) - PrizePicks - Remote - $145k – $165k/yr
- [Software Engineer , Identity](https://hotfix.jobs/jobs/901b31b3-c6b2-42ec-9abf-a5739f55a7ab) - Twilio - Remote - $106k – $156k/yr
- [Software Engineer, Research Infrastructure](https://hotfix.jobs/jobs/aba979c0-7492-4fbb-bbee-841c3aa67496) - Anthropic - San Francisco, CA - $405k – $625k/yr
- [Systems Engineer, Network Protocols & Distributed Systems](https://hotfix.jobs/jobs/19f9e6a4-33f0-4534-9be2-be9b1cbd40b3) - Cloudflare - Austin, TX - €54k – €75k/yr
- [Software Engineer, Platform](https://hotfix.jobs/jobs/e57a78fc-aff8-4ee1-897a-4b82e983e191) - Scale AI - London, United Kingdom

**Apply:** https://hotfix.jobs/jobs/8cccc2f3-6165-4416-8f4c-6d9bed7907e3
**Canonical:** https://hotfix.jobs/jobs/8cccc2f3-6165-4416-8f4c-6d9bed7907e3