Build and maintain reliable data pipelines for healthcare claims data, including ELT/ETL, de-identification, data quality, observability, and integrations. Requires 4+ years experience with Python, SQL, healthcare claims datasets, advanced data modeling, and cloud services.
170k – 190k/yr
Remote4+ YOEData Engineering
About the role
Responsibilities
Build and maintain reliable data pipelines that process raw claims data from diverse sources to enriched, standardized formats using tools including Python, SQL, PostgreSQL, Trino, ClickHouse, Airflow, Datadog, and AWS cloud services.
Build automated observability and monitoring into data quality.
Produce privacy-preserving datasets and protect PHI, implementing de-identification and PII-reduction specifications set with Security, Infrastructure, and Product.
Draft technical design and documentation.
Seek and prioritize technical and product feedback from internal customers.
Iterate quickly with an eye towards value.
Requirements
Bachelor's degree or equivalent experience. Non-traditional backgrounds welcome.
4+ years developing data models, pipelines, and end-to-end analytical solutions.
Programming experience in Python and SQL.
Experience with healthcare claims data, such as EDI 837/835 files, researcher datasets (CMS VRDC, CMS Limited Data Set), commercial claims (Komodo, MarketScan), or equivalents.
Advanced SQL including window functions, subqueries, CTEs, performance tuning and indexes.
Data modeling experience in support of diverse OLAP and OLTP workflows. Whether it’s Kimball or One Big Table, you recognize the tradeoffs and know the rules well enough to break them when it matters.
Comfortable with object-oriented and functional programming patterns, code organization beyond scripts, and debugging workflows.
Experience with ETL/ELT workflows, orchestration (e.g., Airflow), and data engineering patterns (e.g., append-only vs inplace, medallion lakehouse).
Ability to design data systems with scalability, performance, and cost efficiency in mind, particularly for compute and data-intensive workloads.
Software engineering rigor including automated testing, version control, software and data quality.
Thoughtful use of AI coding agents and LLMs in development workflows.
Comfortable working remotely in a collaborative, technical team.
Nice-to-Haves
Revenue cycle and healthcare payments.
Payer-provider contracting.
Experience handling sensitive data in a regulated environment (healthcare, finance, or similar), including practical privacy and de-identification tradeoffs.
Leadership. Act as both a player and a coach to onboard new contributors.
Experience with cloud services (AWS S3, EC2, RDS) and cloud fundamentals (object storage, compute, managed services). Alternative providers are okay too (Azure, GCP).
Compensation and Benefits
Competitive pay with equity options.
Stellar health care plan options (Medical, Dental & Vision), with FSA, DCFSA, & HSA options.
Company-sponsored disability & life insurance.
Unlimited PTO.
401(k) + 4% matching.
Fully remote work + flexible working hours.
$750 work-from-home setup budget.
Paid bi-annual in-person company gatherings.
Quarterly $150 co-hanging stipend to meet up with coworkers.
Senior Data Platform Engineer owning data infrastructure for identity and fraud detection products. Build scalable ETL/ELT pipelines, data observability, and storage layers using Python/Golang, Spark, AWS, and databases. 5+ years experience required; mentor juniors and collaborate with product/DS teams.
170k – 230k/yr
Remote5+ YOEData Engineering
Senior Data Platform Engineer
BeviBoston, MA
Senior individual contributor owning Bevi's full data platform (ingestion through self-service BI). Build and evolve IoT data models, streaming pipelines, governance, observability, and AI integration on Fivetran/Snowflake/dbt/Looker stack. Requires 8+ years owning production data platforms end-to-end.
170k – 210k/yr
On-site8+ YOEData Engineering
Senior Data Engineer
TatariNew York, NY +2
Senior Data Engineer building and owning large-scale ETL pipelines at Tatari to power reporting and measurement products from high-volume TV advertising data. Requires 5+ years experience with Python, SQL, Spark, Airflow, and strong data quality practices in a cross-functional environment.
170k – 200k/yr
Hybrid5+ YOEData Engineering
Senior Data Analytics Engineer
OneSignalNew York, NY +1
Senior Data Analytics Engineer who designs, builds, and maintains scalable data systems and ETL pipelines connecting production data to business tools (Salesforce, Marketo, etc.) on GCP. Combines data engineering with analytics, insights, and ML models to drive sales, marketing, and operational decisions at high-growth SaaS company.
170k – 190k/yr
Remote6+ YOEData Engineering
Lead Data Engineer
CoastNew York, NY
Lead the data engineering team to build and evolve a modern cloud-native data platform, drive data culture, and deliver high-impact data products for risk, customer acquisition, and financial services.