Software Engineer, Backend
Build and operate the infrastructure and platform systems underlying Clay’s products, focusing on scale, reliability, performance, orchestration, and shared developer foundations. The role requires 8+ years of production engineering experience and strong distributed-systems expertise.
About the job
Responsibilities
- Improve scale, reliability, and performance by addressing database contention, memory pressure, throughput ceilings, and concurrency bottlenecks.
- Define performance and reliability targets as load increases.
- Build shared platforms, primitives, contracts, and extension points for product teams.
- Migrate existing consumers onto shared infrastructure.
- Evolve execution and orchestration systems for defining, scheduling, retrying, and observing work.
- Build instrumentation that explains what a run did, where it failed, and why.
- Solve integration problems involving input mapping, data access, credit accounting, and usage accounting.
- Set technical direction and raise engineering standards through design reviews, code reviews, and durable systems.
Requirements
- 8+ years of hands-on engineering experience building and operating production systems at scale.
- Demonstrated expertise in scale, reliability, and performance, or in building internal platforms and frameworks.
- Treat platforms as products and design for internal consumers.
- Own engineering areas end to end, including product judgment and prioritization.
- Communicate nuanced technical ideas clearly across audiences.
- Collaborate effectively to deliver high-quality, understandable, maintainable code.
- Reason about distributed systems, concurrency, queuing, backpressure, failure recovery, and underlying data models.
- Learn unfamiliar technologies quickly.
Technology Stack
- React
- TypeScript
- Python
- Node.js
- AWS Aurora (Postgres)
- Amazon ElastiCache (Redis)
- Amazon ECR
- Amazon ECS (Fargate)
- AWS Lambda
- OpenSearch
- Terraform
- CircleCI
- Netlify
- Playwright
- Amazon CloudWatch
- Datadog
- Mezmo
Nice to Have
- Experience with workflow engines, orchestration systems, or durable execution frameworks.
- Experience turning internal tools into company-wide platforms and migrating consumers.
- Experience with multi-tenant systems that isolate customer workloads.
- Experience running LLM calls or agent loops in production execution paths.
- Experience with search infrastructure, large result sets, or high-volume ingestion.
Compensation and Benefits
- Employees can work with world-class coaches specializing in creativity, management, and related areas.
- The engineering organization works across New York City and San Francisco in two-week sprints, with async planning in Slack, weekly cross-functional meetings, and Linear as the source of truth.
Skills
React, TypeScript, Python, Node.js, Postgres, Redis, AWS, Amazon Ecs, AWS Lambda, Opensearch, Terraform, CircleCI, Playwright, Datadog, Distributed Systems
Similar jobs
Backend Engineering jobsBuild and lead the evolution of SentiLink’s secure, scalable API platform and identity decisioning products. The role requires 6+ years of software development experience, strong Python or Golang skills, and familiarity with PostgreSQL, Docker, and AWS.
Develop system software and integrations for YubiKey production, fulfillment, and manufacturing operations. The role requires 5+ years of software engineering experience, strong native or memory-safe language skills, operating-system knowledge, and familiarity with cryptography.
Designs and develops scalable internal tooling and automation for datacenter infrastructure, including servers, networking, and power and cooling systems. The role requires 5+ years of software development experience, strong distributed-systems expertise, and proficiency in a modern compiled language.
Build and scale distributed network control-plane tooling for edge, backbone, and data center operations at a cloud provider. The role requires 5+ years of backend engineering experience, strong distributed-systems expertise, and proficiency in a modern backend language.
Develop distributed, containerized backend services that process cloud telemetry and real-time events to deliver security insights and recommendations. The role requires 4+ years of scalable systems experience, Go and SQL proficiency, cloud API experience, and familiarity with Kubernetes and Linux.