Member of Technical Staff - Sandbox Platform
Build distributed sandbox infrastructure and developer-facing AI workload platforms across systems, backend, and frontend layers. The role requires strong Rust and Linux systems expertise alongside Python backend and modern web development experience.
About the job
Responsibilities
- Design and implement distributed orchestration infrastructure in Go and Rust.
- Build high-performance networking and coordination components.
- Create infrastructure automation pipelines with Ansible.
- Manage cloud resources and container orchestration.
- Implement scheduling systems for heterogeneous hardware, including CPU, GPU, and TPU.
- Build web interfaces for AI workload management and monitoring.
- Develop REST APIs and backend services in Python.
- Create real-time monitoring and debugging tools.
- Implement user-facing features for resource management and job control.
Requirements
- Systems programming experience with Rust.
- Strong Linux systems knowledge, including networking, namespacing, and performance tuning.
- Virtualization experience with VMs, hypervisors, and low-level resource management.
- Experience with infrastructure automation using Ansible and Terraform.
- Experience with container orchestration using Kubernetes.
- Cloud platform expertise, preferably GCP.
- Experience with observability tools such as Prometheus and Grafana.
- Strong Python backend development experience, including FastAPI and asynchronous programming.
- Modern frontend development experience with TypeScript, React/Next.js, and Tailwind.
- Experience building developer tools and dashboards.
- RESTful API design and implementation experience.
Nice-to-haves
- GPU computing and ML infrastructure experience.
- Knowledge of AI/ML model architecture and training.
- High-performance networking implementation experience.
- Open-source infrastructure contributions.
- WebSocket and real-time systems experience.
Compensation and Benefits
- Cash compensation range of $150,000–$300,000, plus significant equity incentives.
- Flexible work arrangement with the San Francisco office preferred; remote work possible for exceptional candidates.
- Visa sponsorship and relocation support.
- Professional development budget for courses and conferences.
- Team off-sites and conference attendance.
Skills
Go, Rust, Linux, Networking, Ansible, Terraform, Kubernetes, GCP, Prometheus, Grafana, Python, FastAPI, TypeScript, React, Next.js
Similar jobs
Fullstack Engineering jobsBuild scalable full-stack web applications across frontend and backend layers using React, TypeScript, Python, GraphQL, and FastAPI. The role requires at least three years of full-stack development experience and offers hybrid work in San Francisco.
Build scalable web software and infrastructure for a complex, regulated mortgage-servicing platform. The role requires 4+ years of software engineering experience, strong cross-functional communication, and experience with web applications, distributed systems, mobile applications, or infrastructure.
Build and deploy AI agents, automation workflows, and enterprise integrations that transform costly healthcare operations. The role requires production software development, customer collaboration, rapid iteration, and ownership of measurable business outcomes.
Build and own customer-facing data-security features end to end, tackling distributed systems and petabyte-scale production challenges. The role requires 3+ years of software engineering experience, strong backend skills in Go or Python, and expertise in databases and schema design.
Build and ship scalable SaaS features across frontend and backend systems using TypeScript or Python, React, Node.js, and cloud technologies. The role requires 5+ years of experience, strong problem-solving and collaboration skills, and the ability to work across a fast-moving product environment.