Latest DevOps / SRE jobs at Runloop
Job results
Site Reliability Engineer responsible for the reliability, observability, performance, and security of a core AI agent platform including code sandboxes. Requires 5+ years software engineering experience (3+ in SRE/DevOps), strong CS fundamentals, expertise in containers, cloud IaC, monitoring, and distributed systems.
Builds and scales core infrastructure using MicroVMs, Kubernetes, AWS, and GCP to support AI agent development sandboxes. Requires 2+ years experience in infrastructure engineering and systems programming.