Skip to content

Senior Scientific Data Engineer

Leads Scientific Data Engineering team to build data pipelines, schemas, and parsers for pre-clinical lab data using Python/SQL/AI. Architects solutions, mentors juniors, and delivers customer-focused data products with dashboards in React/Streamlit. Requires 8+ years experience.

About the job

Responsibilities

  • Lead the Scientific Data Engineering (SDE) Team and build Tetra Data and productizable solutions.
  • Work with Product Managers and Solution Architects to understand business requirements and build solutions.
  • Take ownership of building data models, prototypes, and integration solutions.
  • Use AI agents to build comprehensive data schemas and parsers for pre-clinical data (R&D lab instruments, manufacturing, CRO, CDMO, ELN, LIMS) with formats: .xlsx, .pdf, .txt, .raw, .fid, vendor binaries.
  • Extract reusable schema components and parsing functions, productize into Python libraries.
  • Build high-quality data pipelines with full unit test and integration test coverage.
  • Build data applications, reports, and dashboards using React, Streamlit, Jupyter notebook.
  • Collaborate with product managers, project managers, business analysts, data architects, and ML engineers.
  • Drive customer value, verify solutions, and act as quality gatekeeper.
  • Lead team-wide process/technology improvements and Agile Sprint commitments.
  • Provide mentorship to junior SDEs.

Requirements

  • 8+ years as Data Engineer or similar.
  • 8+ years in Python and SQL focused on data.
  • Experience leading projects, managing requirements, timelines, and cross-functional customer implementations.
  • Experience with data plotting/dashboarding tools like React and/or Streamlit (strongly preferred).
  • Experience with pre-clinical data and lab scientists (strongly preferred).
  • Excellent communication, attention to detail, and project delivery leadership.

Benefits

  • 100% employer paid benefits for employees and immediate family.
  • 401K.
  • Unlimited PTO.
  • Company paid Life Insurance, LTD/STD.

Skills

Python, SQL, React, Streamlit, Jupyter, AI Agents, Data Pipelines, Unit Testing, Integration Testing, Data Schemas, Data Parsing, Agile, Pre-Clinical Data, Eln, Lims

Addepar

Addepar

United States

Engineering Manager, Reference Data
No salary listedRemote7+ YOEData Engineering

Staff Software Engineer building scalable frameworks for high-performance financial data ingestion/distribution and AI-native products. Requires 7+ years experience with distributed systems, microservices, and data architectures; partners with product teams to drive technical direction.

Alpaca

Alpaca

Remote

Senior Data Engineer
No salary listedRemote5+ YOEData Engineering

Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.

NinjaTrader

NinjaTrader

Chicago, IL

Senior Database Reliability Engineer II
$130k+/yrRemote8+ YOEData Engineering

Leads database architecture, performance, reliability, and developer-tooling initiatives for high-volume trading applications. Requires 8+ years of software engineering experience, expert MySQL skills, backend development expertise, and strong knowledge of distributed systems and database operations.

Deepgram

Deepgram

United States

AI Data Readiness Lead
$165k+/yrRemote5+ YOEData Engineering

Own the company’s metric governance program by defining canonical metrics, enforcing them in semantic and catalog systems, improving data quality, and validating AI-agent outputs. Requires 5+ years in analytics or analytics engineering, strong SQL, production semantic-layer ownership, and experience with AI evaluation and data governance.

Aleph

Aleph

United States

Senior Analytics Engineer
$86k+/yrRemote5+ YOEData Engineering

Own and scale transformation pipelines that convert diverse financial and operational data into reliable FP&A-ready models. The role requires strong SQL and dbt expertise, data integrity and performance skills, and effective collaboration across Engineering and Customer Success.