Latest DevOps / SRE jobs at Invisible Tech
Job results
Provides first-response incident triage and infrastructure stabilization for a production platform in a 24/7 rotation. Requires enterprise experience with Kubernetes, RabbitMQ, PostgreSQL, Azure, production troubleshooting, log-based diagnosis, and calm incident communication.
Own and evolve a platform domain supporting reliable, secure, and cost-effective multi-tenant infrastructure. The role requires strong Kubernetes, cloud, Terraform, Helm, GitOps, Python, security, and agentic coding expertise, along with excellent technical judgment and communication.
Build and operate the Kubernetes-based platform infrastructure, developer tooling, CI/CD systems, and agent infrastructure used across the engineering organization. The role requires production cloud experience, strong infrastructure-as-code skills, Python proficiency, and sound engineering judgment.