Staff / Principal Software Engineer, Platform
Build and operate the platform infrastructure and developer tooling powering Lovable’s AI product, including sandbox runtimes, schedulers, observability, networking, and reliability systems. The role requires 10+ years in platform, infrastructure, or developer experience engineering and onsite work in Stockholm.
About the job
Responsibilities
- Design, build, and maintain core systems enabling the AI product, including a gVisor-based runtime for agentic workloads and a high-throughput sandbox scheduler across multiple cloud providers.
- Bring order and structure to the codebase; integrate or build application frameworks for a growing engineering organization.
- Own and develop the observability stack, from code instrumentation through ingestion to presentation.
- Harden infrastructure against failures, downtime, and slowdowns.
- Plan and implement network infrastructure and cloud strategy.
- Integrate tools for AI-driven development across the engineering organization.
- Identify and drive reliability improvements across engineering teams.
Requirements
- 10+ years of experience in platform, infrastructure, or developer experience engineering, with a track record operating at senior engineering levels.
- Deep experience in production infrastructure, platform engineering, or developer experience, including observability, CI/CD, application frameworks, or productivity tooling.
- Strong proficiency in at least one systems-oriented language such as Go, Rust, or C++.
- Experience designing high-throughput services at scale and/or writing code and tools to support growing engineering organizations in scale-ups.
- Working knowledge of Docker, Kubernetes, and modern infrastructure practices, including schedulers, runtimes, and isolation boundaries.
- Ability to ship high-leverage systems quickly while balancing security, stability, and speed.
- Comfort navigating ambiguity and driving clarity at an organizational level.
- Based in Stockholm or willing to relocate.
- Willingness to work onsite five days per week.
Technology Stack
- Frontend: React, TypeScript
- Backend: Golang, Rust
- Cloud: Cloudflare, Google Cloud, AWS
- Data: ClickHouse, Firestore, Spanner, BigQuery
- DevOps: CI/CD, OpenTelemetry, Grafana, Kubernetes, Terraform
- Local tooling: Nix, DevEnv
Skills
Go, Rust, C++, Docker, Kubernetes, Gvisor, Cloudflare, GCP, AWS, React, TypeScript, ClickHouse, Firestore, Spanner, BigQuery
Similar jobs
DevOps / SRE jobsAs a Principal Operations Engineer, Mechanical, you will be the senior technical authority for mechanical and cooling infrastructure across hyperscale AI data centers. You will lead site assessments, drive operational readiness, review designs, and ensure precision execution of critical systems.
Build and operate high-performance customer compute environments spanning bare metal, Kubernetes, Slurm, GPUs, networking, storage, and observability. The role requires 5+ years of production Linux infrastructure experience and strong expertise in bare-metal Kubernetes, NVIDIA GPUs, virtualization, and networking.
Owns and evolves CI/CD, mobile release, testing, and deployment infrastructure for a production fintech application. The role requires 8+ years in DevOps or related platform disciplines, strong AWS and Kubernetes expertise, and experience with secure mobile release systems.
Leads production reliability for Grafana Cloud’s multi-tenant database products, partnering with product engineering teams to improve SLOs, scalability, observability, automation, and incident response. Requires 8+ years of engineering experience, including substantial SRE or production engineering work, plus strong Kubernetes and cloud expertise.
Designs and operates secure, highly available cloud infrastructure supporting engineering teams, with a focus on GCP, GKE, Terraform, Kubernetes, observability, and developer self-service. Requires 5–8 years of production infrastructure experience and strong cloud, automation, and Linux expertise.