# Staff Data Engineer

**Company:** [Metriport](https://hotfix.jobs/companies/metriport)
**Location:** San Francisco, CA
**Role:** Data Engineering
**Salary:** $200k – $260k/yr
**Experience:** 8+ years
**Skills:** Spark, Parquet, Iceberg, Delta Lake, Snowflake, BigQuery, Redshift, dbt, Airflow, Dagster, Kafka, Kinesis, TypeScript, Python, AWS
**Posted:** 2026-07-30

> Lead the design, architecture, and scaling of Metriport's data platform for ingesting and processing real-time clinical data from millions of patients. Own end-to-end data projects, mentor engineers, support AI/ML workflows, and eventually lead a team while staying hands-on. Requires 8+ years building large-scale data platforms with modern cloud-native tools.

## Job Description

## What you'll be doing
We ingest clinical data for millions of patients from external healthcare sources, with continuous updates for a growing subset of those patients. You'll own the architecture and evolution of the data platform that powers our product — and ship it to customers fast.

Day to day, that looks like:
- Setting the technical direction for our data platform: evolving our warehouse, data lake, and ETL/ELT architecture to scale with patient and customer growth, and picking the right tools (batch and streaming processing, table formats, orchestration, query engines).
- Driving the critical data projects end-to-end: writing Design Documents, shipping v0's quickly, and iterating to v1 and beyond.
- Supporting AI/ML efforts: making sure the AI Engineers have the data they need.
- Multiplying the team: mentoring engineers on data fundamentals, reviewing designs and PRs, and judging when to invest in quality vs. ship fast.
- Eventually, acting as Team Lead for a group of engineers: breaking down and delegating work, unblocking teammates, and owning your team's delivery — while staying hands-on in the code.
- Driving bi-weekly sprint planning and retros, contributing to the engineering roadmap, joining our daily 30-min remote stand-up at 7:30am PST (our only mandatory meeting), and taking part in the on-call rotation.

Example projects you could own:
- Rearchitecting our patient data consolidation pipeline (deduplication, normalization, hydration) to handle 100x today's volume without 100x the cost.
- Building pipelines that deliver clinical data directly into customers' data warehouses, reliably and at scale.
- Designing the ingestion path for customers pushing large volumes of their own data into the platform.
- Building document-processing pipelines that extract structured data from PDFs, images, and free text to feed ML models.

## Requirements
- 8+ years of engineering experience, with significant depth building, operating, and scaling data platforms processing terabytes of data and millions-to-billions of events a day.
- You've designed data architectures end-to-end — ingestion, storage, processing, warehousing, serving — and owned the tradeoffs (cost, latency, correctness, operability) at each layer.
- Deep experience with modern, cloud-native data stacks: e.g., Spark, open table formats (Parquet, Iceberg, Delta) on S3, warehouses (Snowflake, BigQuery, Redshift), dbt, orchestration (Airflow, Dagster), and streaming (Kafka, Kinesis). Breadth matters — you'll be picking our stack.
- Strong software engineering fundamentals — you write production code (we're a TypeScript shop, with Python in data/ML workflows), not just orchestration configs.
- Experience mentoring or guiding other engineers — through code reviews, pairing, design feedback, or onboarding.
- Located in San Francisco / Bay Area, or willing to relocate.

## Bonus
- Experience leading engineers.
- Experience building or supporting ML/data science workflows (feature pipelines, model inputs/outputs, unstructured data extraction).
- Healthcare standards/technologies: FHIR, HIE, IHE, EHR/EMR, NPI, TEFCA, ADT, HL7, HEDIS, RAF, SNOMED, LOINC, ICD-10, etc.

## Benefits
- Competitive equity + compensation package
- Full family Platinum health insurance, dental, and vision coverage
- 401(k) retirement plan + matching
- Flexible work from home or in-office
- Healthy lunches are complimentary when working in-office (and breakfast + dinners as needed)
- Quarterly company off-sites with the team
- MacBook provided by us
- Unlimited PTO (we work hard, but trust you to take time you need to be at your best)

## Similar jobs

- [Staff Software Engineer, Communication & Connectivity](https://hotfix.jobs/jobs/db2a925e-c86f-4b1c-959e-0319e66571e8) - Airbnb - Remote - $204k – $255k/yr
- [Staff Data Engineer, Market Data](https://hotfix.jobs/jobs/f38156de-5707-4a30-9f71-afbc40d76fae) - Coinbase - Remote - $207k – $244k/yr
- [Staff Software Engineer, Ads Measurement & Orchestration](https://hotfix.jobs/jobs/a8a95b11-6cd7-494a-beff-d80cc0f57bd5) - Databricks - New York, NY - $191k – $254k/yr
- [Staff Software Engineer, Data Warehouse](https://hotfix.jobs/jobs/25810bb4-ca57-40e6-abce-6cd6d9df11fb) - Commure - Mountain View, CA - $210k – $275k/yr
- [Senior/Staff Software Engineer, Data Platform](https://hotfix.jobs/jobs/b65925ec-a9ae-4ecc-9504-9ffe994733f7) - Axion - San Francisco, CA - $210k – $265k/yr

**Apply:** https://hotfix.jobs/jobs/995ebeb2-7fc0-450e-8805-7caee5840b6c
**Canonical:** https://hotfix.jobs/jobs/995ebeb2-7fc0-450e-8805-7caee5840b6c