Software Engineer IV, Datastore
Leads architecture and development of Beacon’s backend Datastore, including PostgreSQL models, Kafka pipelines, and GraphQL APIs for clinical and scientific data. Requires 7+ years of backend experience, strong SQL and data-platform expertise, and proficiency in Julia or Python.
About the job
Responsibilities
- Lead design and architecture for complex Datastore systems, weighing implementation trade-offs with data-driven engineering principles.
- Design and scale backend data infrastructure, including PostgreSQL data models, Kafka-driven event pipelines, and GraphQL APIs.
- Own pipelines that deliver dataset snapshots into the data warehouse.
- Translate product, scientific, clinical, and partner requirements into integrations, services, and scalable data-platform solutions.
- Debug and profile cross-system problems, improving structure, test coverage, tooling, operational robustness, documentation, and dashboards.
- Draft RFCs and lead technical discovery and design phases.
- Plan individual and team work while balancing speed, generalizability, and technical debt.
- Mentor engineers, provide actionable feedback, and serve as a technical lead or engineering-practice coordinator.
- Collaborate across platform, application, scientific, clinical, and support teams.
Requirements
- 7+ years of backend development experience, including at least 3 years focused on data platforms and backend infrastructure for data-centric applications.
- Advanced expertise in at least one of data modeling, distributed systems, API design, or database engineering.
- Strong SQL proficiency and production PostgreSQL experience, including schema design, indexing, query tuning, and query plans.
- Hands-on experience with data streaming and event processing using Kafka or an equivalent technology.
- Proficiency in Julia or Python.
- Experience with GraphQL and TypeScript/JavaScript, including Node.js.
- Experience deploying and operating services in containerized environments.
- Familiarity with Kubernetes and infrastructure-as-code tools such as Terraform or Helm.
- Track record of mentoring engineers and improving team practices, tooling, testing, or documentation.
- Strong written and verbal communication skills.
- Ability to work effectively in hybrid or fully remote teams.
- Experience with or interest in LLM-assisted or agentic coding tools in production, with appropriate guardrails.
Nice-to-haves
- Julia experience.
- Experience with Kafka, RabbitMQ, Pulsar, Amazon Kinesis, or comparable cloud event-streaming platforms.
- Interest in data infrastructure for healthcare and scientific applications.
Skills
Postgres, SQL, Kafka, GraphQL, Julia, Python, TypeScript, JavaScript, Node.js, Kubernetes, Terraform, Helm, Distributed Systems, Data Modeling, Llm Coding Tools
Similar jobs
Backend Engineering jobsBuild and operate high-throughput blockchain infrastructure, APIs, and platform primitives integrating protocols such as Ethereum and Bitcoin with internal services. Requires 5+ years of software engineering experience, distributed-systems expertise, and hands-on crypto infrastructure experience.
Design, build, and operate Cloudflare’s globally distributed cache and reverse-proxy data plane, improving performance, correctness, and resilience across the edge. Requires at least 4 years of production systems experience and proficiency in a systems or backend language.
Senior backend engineer designing and operating reliable billing and financial systems, APIs, data models, and distributed workflows. The role requires 5+ years of professional software development experience, strong backend expertise, and collaboration across Product, Finance, Operations, and Data.
Senior individual contributor responsible for designing, building, operating, and improving large-scale backend services, APIs, and telemetry pipelines in Go and Python. The role requires production systems ownership, distributed-systems expertise, incident leadership, mentoring, and technical design leadership.
Build and operate backend services, data pipelines, storage, and retrieval systems that provide trusted context to agentic platforms and product applications. The role requires 8+ years of software engineering experience, distributed-systems expertise, cloud infrastructure knowledge, and strong data modeling skills.