Scientific Data Engineer
Builds ETL pipelines and Python scripts to ingest, transform, and manage customer R&D datasets for a scientific web platform. Targets recent graduates with Python scripting, data manipulation experience, and CS coursework.
About the job
Responsibilities
- Structure and ingest customer data for the Uncountable Web Platform.
- Manipulate, transform, and upload R&D data by writing Python scripts.
- Setup ETL pipelines between Uncountable and other data warehouses.
- Provide new solutions for users to bring in and export their data.
Requirements
- Computer science, data analysis and/or data engineering coursework.
- Interest in a software engineering career and technology startups.
- Affinity and familiarity working with data in Excel.
- Scripting or coding experience in Python to efficiently manipulate large volumes of data.
Preferred Qualifications
- Knowledge of SQL, experience writing queries and working with databases.
Benefits
- Competitive Salary and Equity (ISO Grant).
- Health and Dental Insurance.
- 401K with Employer Contribution.
- 17 days PTO annually.
Skills
Python, SQL, ETL, Data Warehousing, Excel, Data Processing, Data Ingestion
Similar jobs
Data Engineering jobsAnalytics Engineering intern building dimensional data models, SQL pipelines, quality controls, and self-serve datasets or dashboards. Requires current quantitative-degree study, SQL proficiency, programming familiarity—preferably Python—and clear technical communication.
Data Engineering Intern supporting scalable pipelines and infrastructure for analytics and machine learning workloads. Requires Python and SQL proficiency, cloud familiarity, and exposure to modern software architecture or AI/API integrations.
Build and optimize scalable data pipelines, reusable datasets, and federated data quality systems for healthcare analytics. The role requires at least 2 years of data or software engineering experience and strong Python, SQL, AWS, orchestration, database, and warehouse expertise.
Build and operate production data pipelines and transformation layers that turn heterogeneous business, identity, and fraud data into reliable inputs for entity resolution, scoring, and customer APIs. The role requires at least one year of data engineering experience with Python, SQL, cloud platforms, and modern pipeline tooling.
Build and scale secure, cloud-native data pipelines and orchestration systems for healthcare imaging, biomarkers, analytics, and AI. The role requires Python, SQL, Airflow or similar orchestration, cloud platforms, Databricks, and distributed processing experience.