Software Engineer, Platform
Owns and scales platform infrastructure including edge/cloud services on Cloudflare, GCP, Vercel and data layers like Spanner, ClickHouse, Postgres to serve millions of LLM requests daily. Requires 5+ years in production infrastructure with cloud platforms, databases, and full-stack TypeScript expertise.
About the job
What You'll Do
- Own and evolve our edge and cloud infrastructure across Cloudflare, Google Cloud, and Vercel.
- Scale and operate our data layer including Spanner, ClickHouse, and Postgres.
- Ensure we are optimizing for performance when serving LLM inference as traffic rapidly grows.
- Partner with engineering leadership on capacity, reliability, and cost across the routing layer, with ownership of the systems carrying production traffic.
- Set the bar and playbook for how we run infrastructure and operations as the team grows — tooling, observability, on-call, and the patterns other engineers build against.
About You
- 5+ years building and operating production infrastructure at companies where uptime, latency, and cost matter.
- Proven experience with cloud platforms (GCP, AWS, Azure) and edge-first serverless platforms (e.g. Cloudflare Workers).
- Deep expertise in operating large scale databases (e.g Postgres, Spanner, etc).
- A full-stack TypeScript shop won't faze you; you can move across the stack when the platform needs it.
- High agency and a bias toward action. You don't wait for tickets — you see the bottleneck and fix it.
- AI-forward in your workflow. You use coding agents, MCPs, and LLMs heavily and have opinions about what works.
- Pragmatic about tradeoffs between speed and simplicity.
Bonus Points
- Existing user of OpenRouter, or active side projects in AI products/infrastructure or developer tooling.
Compensation: Base salary $215,000 - $285,000 plus benefits & equity (US full-time). International compensation varies by local market.
Skills
GCP, Cloudflare Workers, Spanner, ClickHouse, Postgres, TypeScript, AWS, Azure, Llm Inference
Similar jobs
DevOps / SRE jobsBuild and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Build and own production-grade AI agent infrastructure across multiple clouds, with responsibility for Kubernetes, Terraform, observability, security, reliability, and automation. Requires 5+ years of cloud infrastructure experience and strong CI/CD, networking, and production operations expertise.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.
Own Mercor’s internal identity and cloud platform infrastructure as code, automating provisioning, access management, secrets, and employee lifecycle workflows. The role requires production Terraform, Okta, SCIM, and multi-cloud IAM experience, plus strong automation, incident response, and documentation skills.
Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.