Senior Data Engineer
Senior Data Engineer responsible for designing and deploying scalable data infrastructure, orchestration models, and analytics tooling to enable data-driven decisions, ML products, and enterprise reporting at Vanta. Requires 4+ years data experience, software engineering mindset, modern data stack proficiency, and passion for secure, compliant data systems.
About the job
What you’ll do as a Senior Data Engineer at Vanta
- Design and deploy data infrastructure needed to drive data-driven decision-making solutions
- Design and implement complex data orchestration models, modeling metadata, scaling reporting tools for data science and ML products users
- Be the company’s expert on data administration, data management and scalable data systems
- Write highly tuned, scalable SQL queries running over large-scale, heterogeneous data warehouses
- Work with the Product and Enterprise Engineering system teams to structure source systems for reporting consumption across the enterprise
- Help maintain CDC pipelines to power customer reporting
- Help develop front end applications to expose analytical data sets enterprise wide
How to be successful in this role
- Have at least four years of experience working with data and two years of experience in Software Engineering or a related field
- Have experience with common analytics tooling (e.g. Stitch/Fivetran, Snowflake/BigQuery/Redshift, dbt, Airflow, Dagster)
- Have good working knowledge of AWS data infra systems and Terraform
- Bring a system-oriented and software engineering mindset to the Data Engineering practice. We’re looking to build frameworks that manage data, and minimize bespoke queries
- Deep knowledge of crafting dimensional and fact models in modern data fashion
- Have a passion for enabling the developer experience of data, and being obsessed with giving data super powers across the company
- Desire to lead the industry in security, anonymization, and compliance management when it comes to data warehousing
- Open to using AI to amplify their skills and strengthen their work - demonstrating curiosity, a willingness to learn, and sound judgment in applying AI responsibly to improve efficiency and impact
What you can expect as a Vanta’n
- Industry-competitive salary and equity
- Comprehensive medical, dental, and vision coverage, with 100% of employee-only benefit premiums covered for most medical plans
- 16 weeks fully-paid Parental Leave for all new parents
- Health & wellness stipend
- Remote workspace, internet, and cellphone stipend
- Commuter benefits for team members who report to the SF and NYC office
- Family planning benefits
- Matching 401(k) contribution with immediate vesting
- Flexible PTO policy, plus 80 hours of Sick Time
- 11 company-paid holidays
- Virtual team building activities, lunch and learns, and other company-wide events!
Skills
SQL, Snowflake, BigQuery, Redshift, dbt, Airflow, Dagster, AWS, Terraform, Fivetran, Stitch, Cdc, Dimensional Modeling, Fact Modeling
Similar jobs
Data Engineering jobsBuild and operate Jump’s Snowflake and dbt data platform, including ELT pipelines, governance, testing, monitoring, and analytics models. The role requires 6+ years of data engineering experience, production-grade SQL and Snowflake expertise, deep dbt knowledge, and strong Python and software engineering practices.
Build and own large-scale data models, batch and real-time pipelines, and data infrastructure that provide reliable datasets and insights across Plaid. The role requires 4+ years of data engineering experience, strong SQL and Python skills, and expertise with modern warehouses, lakes, and orchestration tools.
Builds and owns scalable SQL/Python data pipelines, golden datasets, and workflows using DBT, Airflow, Redshift for large-scale data (500TB+). Collaborates cross-functionally to enable data-driven decisions at Plaid. Requires 4+ years data engineering experience.
Build and operate low-latency systems that capture, normalize, and distribute real-time market data for institutional trading. The role requires backend engineering experience, Java or C++, market data infrastructure knowledge, and exchange connectivity expertise.
Build and operate large-scale revenue data pipelines powering billing and cost attribution, while improving reliability, latency, and correctness. The role requires strong Spark and Airflow experience, cross-functional problem-solving, and operational ownership of mission-critical production systems.