Data Engineer
Designs and owns mission-critical data pipelines to enable decision-making across data science, growth, sales, marketing, and product teams. Requires 5+ years experience with scalable pipelines (preferably Airflow), Python, and advanced SQL.
About the job
What you'll do
- Work between our engineering organization and stakeholders from our data science, growth, sales, marketing, and product teams, to understand the data needs of the business and produce pipelines, data marts, and other data solutions that enable better product and growth decision-making.
- Design and update our foundational business tables in order to simplify analysis across the entire company.
- Continue to improve the performance and reliability of our data warehouse.
- Build and enforce a pattern language across our data stack, ensuring that our data pipelines and tables are consistent, accurate, and well-understood.
Who you are
- You have 5+ years of professional experience designing, creating and maintaining scalable data pipelines, preferably in Airflow.
- You've wrangled enough data to understand how often the complex systems that produce data can go wrong.
- You are proficient in at least one programming language (preferably Python), and are willing to become effective in others as needed to get your job done.
- You are highly effective with SQL and understand how to write and tune complex queries.
- You're passionate and thoughtful about building systems that enhance human understanding.
- You communicate with clarity and precision in written form; experience communicating with graphs and plots.
Skills
Airflow, Python, SQL, Data Pipelines, Data Warehouse, Apache Airflow
Similar jobs
Data Engineering jobsBuild scalable data pipelines, infrastructure, and quantitative models that support experimentation, forecasting, and business decision-making. The role requires 4+ years of production data engineering experience, strong Python and SQL skills, distributed computing expertise, and a quantitative degree.
Own the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.
Builds and optimizes scalable data pipelines, storage, and OLAP databases for ML training, analytics, and product features. Requires 5+ years in data engineering, proficiency in Python/SQL/cloud platforms, and distributed systems experience.
Build and operate scalable data infrastructure, including partner data sharing, identity graph foundations, and governed batch and real-time platforms. The role requires 5+ years of data, distributed systems, infrastructure, or backend engineering experience and strong cloud and data-platform expertise.
Builds scalable data pipelines and data engine architecture for machine learning, integrating foundation models to automate labeling and discovery. The role requires 5+ years of experience, modern ML infrastructure expertise, and U.S. citizenship with security-clearance eligibility.