Data Engineer
Build and scale reliable data pipelines for the Growth team, establish dbt and data-quality standards, and enable self-service analytics across cross-functional teams. The role requires strong data engineering experience, expert dbt knowledge, and proficiency in Python and SQL.
About the job
Responsibilities
- Own and streamline data pipelines from ingestion to delivery, ensuring they are reliable, scalable, and efficient.
- Implement and maintain dbt best practices, data standards, and quality controls.
- Build self-service tooling and write clear documentation to empower cross-functional teams, including Engineering, Revenue Operations, and Growth, to independently meet their data needs.
- Create AI agents to serve data questions for others.
Requirements
- Demonstrated experience formalizing and scaling data pipelines, ideally as one of the first or early data engineers on a growing team.
- Expert knowledge of dbt.
- Proficiency with tools across the modern data stack, including Python, SQL, and BI tools.
- Structured thinking and the ability to simplify complex, messy data into clear structures.
- Strong ability to understand varied stakeholder requirements and build generalized solutions.
Benefits
- Annual discretionary learning and development stipend.
- Annual discretionary social travel stipend.
- Annual company offsite.
- Monthly coworking stipend for employees not located near a main hub.
Skills
dbt, Python, SQL, BI Tools, Data Pipelines, Data Quality, Data Modeling, AI Agents
Similar jobs
Data Engineering jobsBuild and scale distributed data platforms, database systems, delivery services, and APIs, with emphasis on reliability, performance, observability, and data integrity. Requires 3+ years of software development experience with distributed systems and databases; Golang experience is preferred.
Build the data infrastructure, processing pipelines, curation strategies, and tooling that power frontier AI model training. The role requires strong distributed-data engineering experience and the ability to measure how data quality affects model outcomes.
Build and operate distributed web-crawling systems that source, extract, evaluate, and prepare large-scale web data for frontier AI models. The role requires hands-on crawler or scraping experience and strong distributed-systems engineering skills.
Own the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.
Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.