Skip to content

Data Engineer

Build and scale secure, cloud-native data pipelines and orchestration systems for healthcare imaging, biomarkers, analytics, and AI. The role requires Python, SQL, Airflow or similar orchestration, cloud platforms, Databricks, and distributed processing experience.

About the job

Responsibilities

  • Design, build, and operate scalable data pipelines and orchestration systems for imaging and biomarker data.
  • Build secure, compliant data infrastructure aligned with healthcare standards such as HIPAA and GDPR.
  • Improve performance and fault tolerance across distributed data workflows.
  • Integrate DICOM and other data pipelines into the broader data lake and analytics stack.
  • Implement data validation, observability, testing, and monitoring for production pipelines.
  • Collaborate with data science, machine learning, AI, and product teams to ensure data reliability and accessibility.
  • Automate data ingestion, transformation, cataloging, and access for analytics and machine learning use cases.

Requirements

  • 1–2+ years of experience in data engineering, ETL development, or cloud-based data orchestration.
  • Strong proficiency in Python and SQL.
  • Hands-on experience with workflow orchestration tools such as Airflow or Prefect.
  • Experience with cloud platforms and services, especially AWS; Azure experience is also relevant.
  • Experience building on Databricks and using distributed processing frameworks such as Apache Spark or Dask.
  • Understanding of data validation, observability, testing, and production best practices.
  • Strong communication and documentation skills, with the ability to work across technical teams.

Nice-to-haves

  • Experience in healthcare, imaging, biotech, or other high-trust data environments.
  • Experience working with PHI-sensitive data and healthcare standards.
  • GCP familiarity.
  • 1–2+ years of experience in data science, analytics, or applied research if holding a PhD or having published research.

Benefits

  • Stock options.
  • Comprehensive health, dental, and vision plans for employees and families.
  • Wellness and commuter benefits.
  • Competitive vacation policy.
  • Flexible working hours and a learning-focused culture.

Skills

Python, SQL, Apache Airflow, Prefect, AWS, Azure, Databricks, Spark, Dask, ETL, Dicom, HIPAA, GDPR, Data Observability, Data Validation

Coinbase

Coinbase

San Francisco, CA

Analytics Engineer Intern
$50+/hrHybridData Engineering

Analytics Engineering intern building dimensional data models, SQL pipelines, quality controls, and self-serve datasets or dashboards. Requires current quantitative-degree study, SQL proficiency, programming familiarity—preferably Python—and clear technical communication.

Coinbase

Coinbase

San Francisco, CA

Data Engineer Intern
$50+/hrHybridData Engineering

Data Engineering Intern supporting scalable pipelines and infrastructure for analytics and machine learning workloads. Requires Python and SQL proficiency, cloud familiarity, and exposure to modern software architecture or AI/API integrations.

Garner Health

Garner Health

New York, NY

Data Engineer III
$166k+/yrHybrid2+ YOEData Engineering

Build and optimize scalable data pipelines, reusable datasets, and federated data quality systems for healthcare analytics. The role requires at least 2 years of data or software engineering experience and strong Python, SQL, AWS, orchestration, database, and warehouse expertise.

Baselayer

Baselayer

San Francisco, CA

Data Engineer
$120k+/yrHybrid1+ YOEData Engineering

Build and operate production data pipelines and transformation layers that turn heterogeneous business, identity, and fraud data into reliable inputs for entity resolution, scoring, and customer APIs. The role requires at least one year of data engineering experience with Python, SQL, cloud platforms, and modern pipeline tooling.

Navan

Navan

Dallas, TX

Data Engineer
No salary listedHybrid2+ YOEData Engineering

Build analytics data models, pipelines, dashboards, and AI-enabled automation for product and operational decision-making. The role requires 2–3+ years of data engineering or analytics experience, advanced SQL, dbt, cloud data warehouse, Python, and BI expertise.