Staff Software Engineer, Platform
Defines the technical direction and builds the secure, reliable, cloud-agnostic platform enabling enterprise customers to deploy Lovable across major cloud providers. The role requires Staff-level infrastructure expertise, strong programming skills, distributed-systems experience, and architectural ownership.
About the job
Responsibilities
- Define the architecture for enterprise deployments across AWS, Azure, and Google Cloud.
- Build cloud-agnostic systems that give customers flexibility in how they deploy Lovable.
- Design secure sandboxing and runtime environments.
- Set the technical bar for networking, encryption, identity, and enterprise infrastructure.
- Build the internal platform that enables engineering teams to ship faster.
- Mentor engineers and raise the engineering bar across the organization.
- Partner directly with enterprise customers and cloud providers on complex infrastructure challenges.
- Build software, make architectural decisions, and solve complex infrastructure problems with customers.
Requirements
- Experience architecting platform or infrastructure systems that have scaled in production.
- Experience building enterprise products and working with enterprise customers.
- Deep expertise across AWS, Azure, or Google Cloud.
- Experience designing distributed systems used across multiple teams.
- Strong programming skills in Go, Rust, Java, C++, or similar.
- Experience building secure, reliable systems operating under demanding production workloads.
- Experience operating at Staff, Principal, founding engineer, or CTO-level scope.
- A product mindset and focus on the people using the systems you build.
- Strong platform and infrastructure experience with a track record of setting technical direction.
Technical Stack
- Languages: Go, Rust, TypeScript
- Cloud: AWS, Azure, Google Cloud, Cloudflare
- Infrastructure: Kubernetes, Terraform, Temporal, OTEL, CI/CD
- Data: PostgreSQL, ClickHouse, Firestore, Spanner, BigQuery
Application Information
- Applications should be submitted in English.
Skills
Go, Rust, TypeScript, Java, C++, AWS, Azure, GCP, Cloudflare, Kubernetes, Terraform, Temporal, OpenTelemetry, Postgres, Distributed Systems
Similar jobs
DevOps / SRE jobsBuild and operate a Kubernetes-native control plane for provisioning, scheduling, self-healing, and optimizing GPU inference infrastructure. The role requires strong software engineering, durable workflow orchestration, reconciliation systems, event-driven architecture, and platform API experience.
Build and operate foundational observability infrastructure spanning telemetry pipelines, profiling, tracing, and diagnostic tooling across large-scale compute clusters. The role requires deep systems-level experience and 10+ years of relevant industry experience.
Build and operate a Kubernetes-native control plane for provisioning, scheduling, self-healing, and optimizing GPU inference infrastructure. The role requires strong software engineering, durable workflow orchestration, reconciliation systems, event-driven architecture, and platform API experience.
Leads reliability engineering for critical AI serving systems, spanning SLOs, observability, high availability, and incident response. Requires strong distributed-systems or infrastructure experience, with model-serving, accelerator, networking, and resilience-testing expertise valued.
Staff Engineer responsible for deploying, integrating, maintaining, and developing an AI training factory across isolated environments. The role requires 7+ years of related experience, cloud and Kubernetes expertise, Linux networking knowledge, application support skills, and automation experience.