Distributed Systems Engineer - Platform
Builds scalable distributed systems including queueing, state stores, and execution layers for developer tools platform. Requires experience with Go, distributed systems at scale, and strong engineering judgment. Works with US PST overlap.
About the job
Responsibilities
- Architect and implement solutions in our queueing layer, state store, and execution layer (e.g., concurrency over time, function debounce)
- Plan and implement improvements on throughput and latency at hundreds of thousands to millions of requests per second
- Contribute to systems architecture and infrastructure changes as we grow
- Collaborate with team members to expose internal data across metrics stores, APIs, and customer dashboards
- Work with backend engineers to design APIs for Inngest cloud dashboard, dev server, and CLIs
- Dogfood the Inngest product and develop ideas for improvements, features, or integrations
- Communicate with users through Github, email, and Discord
- Write technical specs for features and documentation for users
Requirements
- Experience working on distributed systems for several years
- Professional experience with Go or similar statically typed languages for two years or more
- Architected or designed systems that handle scale
- Understand engineering trade-offs and make correct judgement calls
- Understand how to observe, monitor, and maintain systems
- Appreciate simplicity in design and build
Nice-to-Haves
- Work with compliance (SOC2, ISO27001, HIPAA)
- Experience or understanding of networking
- Experience managing and maintaining systems (e.g., SRE roles)
Tech Stack
- Backend: Go, Postgres, Redis, Clickhouse, PubSub/Kafka, Kubernetes
- APIs: gRPC (internal), GraphQL, REST
- Hosted on: AWS, GCP, Bare Metal
- Tools: Github, Linear, Slack, Notion, Figma
Skills
Go, Kubernetes, Postgres, Redis, Kafka, gRPC, GraphQL, AWS, GCP, ClickHouse
Similar jobs
DevOps / SRE jobsSummer 2027 internship on a Site Reliability Engineering team, building software and automation for deployment, operations, monitoring, and reliability. Requires a software engineering foundation, programming experience, and strong problem-solving and collaboration skills.
Customer-facing DevOps Engineer helping organizations implement secure, compliant cloud infrastructure through the DuploCloud platform. Requires 2–3 years of cloud or DevOps experience, containerization expertise, public cloud knowledge, and strong customer communication skills.
Build and operate large-scale scheduling, storage, caching, and networking infrastructure for AI training and inference. The role targets PhD researchers graduating by December 2026 with systems research depth and strong programming and performance-measurement skills.
Infrastructure and site reliability intern building and operating on-premises backend infrastructure for a semiconductor fabrication environment. The role emphasizes systems programming, Linux, networking, reliability, observability, automation, and performance engineering.
Supports cloud infrastructure, automation, CI/CD, monitoring, and service reliability while learning alongside a global DevOps team. The entry-level role requires a bachelor’s degree, foundational systems knowledge, and exposure to cloud and DevOps tools.