Data Engineer Intern
Data Engineering Intern supporting scalable pipelines and infrastructure for analytics and machine learning workloads. Requires Python and SQL proficiency, cloud familiarity, and exposure to modern software architecture or AI/API integrations.
About the job
Responsibilities
- Build and maintain data pipelines that ensure data integrity and accessibility for analytics and machine learning workloads.
- Support scalable data infrastructure using Python and cloud-based technologies.
- Integrate and manage data from internal systems and external APIs.
- Contribute to analytical deep dives with senior team members, translating findings into documentation and recommendations.
- Apply data engineering best practices, including data governance, security, and performance optimization.
Requirements
- Completed or in-progress coursework in Computer Science, Data Engineering, or equivalent experience with demonstrated programming proficiency in Python, including data structures and algorithms.
- Hands-on experience writing SQL for data analysis and working with relational database concepts.
- Familiarity with cloud platforms such as AWS, Google Cloud, or Azure.
- Demonstrated interest in AI, large language models, or API integrations.
- Exposure to software architecture concepts such as microservices, distributed systems, or cloud-native design through coursework or projects.
- Responsible use of generative AI with human oversight to deliver business-ready outputs and improve workflow efficiency, cost, and quality.
Compensation
- Hourly rate: $50 USD.
Skills
Python, Data Structures, Algorithms, SQL, Relational Databases, AWS, GCP, Azure, APIs, Microservices, Distributed Systems, Generative AI
Similar jobs
Data Engineering jobsAnalytics Engineering intern building dimensional data models, SQL pipelines, quality controls, and self-serve datasets or dashboards. Requires current quantitative-degree study, SQL proficiency, programming familiarity—preferably Python—and clear technical communication.
Build and maintain reliable data pipelines, warehouses, and lightweight data applications supporting analytics, operational models, and AI/ML workflows. The role requires SQL, Python, ETL, and data warehousing knowledge, with 1+ year of relevant experience preferred.
Builds customer intelligence workflows that turn product usage, engagement, and contract data into actionable insights for Customer Success. The role requires strong Python and SQL skills, a bachelor's degree, and an interest in applied AI, automation, and predictive customer health modeling.
Build and operate production data pipelines and transformation layers that turn heterogeneous business, identity, and fraud data into reliable inputs for entity resolution, scoring, and customer APIs. The role requires at least one year of data engineering experience with Python, SQL, cloud platforms, and modern pipeline tooling.
Build and operate the streaming, storage, query, and self-service infrastructure underlying the company’s data platform. The role suits an early-career engineer with 1–3 years of experience, a computer science bachelor’s degree, programming skills, and interest in distributed systems.