Data Engineer
Own data and analytics end-to-end: architect internal systems, build metrics/dashboards, and translate customer and product signals into structured inputs for AI agents.
About the job
What You’ll Do
- Own internal metrics, dashboards, and reporting across product, GTM, and usage
- Build systems to track, analyze, and interpret key signals across the customer and product journey
- Help design and implement internal “sensors” to uncover buried insights
- Translate messy, real-world data into clean, structured inputs for our reasoning agents
- Define and evolve our analytics stack — dbt, Snowflake, Mode, etc. — with an eye toward speed and clarity
- Work cross-functionally with founders, engineering, and forward-deployed teams to prioritize and act on what matters
Who You Are
- A builder — you like owning problems end-to-end and working from first principles
- Comfortable with both SQL and Python, and fluent in the modern data stack
- Analytical by instinct — you’ve spent enough time with data to develop intuition and taste
- Bonus: background in math, physics, or quant research
- Bonus: experience working with GTM teams or building analytics for sales/product workflows
Skills
SQL, Python, dbt, Snowflake, Mode
Similar jobs
Data Engineering jobsBuild scalable data pipelines, infrastructure, and quantitative models that support experimentation, forecasting, and business decision-making. The role requires 4+ years of production data engineering experience, strong Python and SQL skills, distributed computing expertise, and a quantitative degree.
Own the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.
Builds and optimizes scalable data pipelines, storage, and OLAP databases for ML training, analytics, and product features. Requires 5+ years in data engineering, proficiency in Python/SQL/cloud platforms, and distributed systems experience.
Build and operate scalable data infrastructure, including partner data sharing, identity graph foundations, and governed batch and real-time platforms. The role requires 5+ years of data, distributed systems, infrastructure, or backend engineering experience and strong cloud and data-platform expertise.
Builds scalable data pipelines and data engine architecture for machine learning, integrating foundation models to automate labeling and discovery. The role requires 5+ years of experience, modern ML infrastructure expertise, and U.S. citizenship with security-clearance eligibility.