Member of Technical Staff - Platform Eng
Builds and scales platform infrastructure enabling AI agents to integrate with tools like GitHub and Salesforce. Evolves APIs, manages code runtimes, optimizes performance, and collaborates with product teams on backend distributed systems.
About the job
Responsibilities
- Evolve platform primitives and APIs: auth, automatic refreshes, triggers, tool search, tool planning, orchestrating and managing sandboxes.
- Manage multiple runtimes for code execution across lambdas and firecracker.
- Performance optimization: tracing, CPU/heap profiling, profiling, optimizing db queries, optimizing temporal workflows.
- Work closely with product engineering teams and customers to win their workloads, improving the product in the process.
- Writing articulate docs.
Requirements
Must haves:
- Core platform engineering.
- Worked extensively in scaling backend distributed systems.
- Know how to maintain reliable systems while shipping fast.
- Able to hold many different parts of the platform in mind at once.
- AI native: built with language models, built for language models.
- Linux: comfortable in a Linux environment.
- Typist: can write docs well and explain complex ideas clearly.
- Human: build trust and admit what you don’t know.
Optional:
- Multiple years of experience writing TypeScript, Go.
- Contributions to a major open source project.
- Started companies or built large side projects.
Skills
Distributed Systems, Linux, TypeScript, Go, AWS Lambda, Firecracker, Temporal, Language Models, APIs, Tracing
Similar jobs
DevOps / SRE jobsBuild and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.
Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.