Analytics Engineer - X
Build scalable data pipelines, infrastructure, and quantitative models that support experimentation, forecasting, and business decision-making. The role requires 4+ years of production data engineering experience, strong Python and SQL skills, distributed computing expertise, and a quantitative degree.
About the job
Responsibilities
- Design, implement, and optimize end-to-end data pipelines for high-volume datasets using Spark, Kafka, Flink, and related technologies.
- Develop quantitative models and statistical frameworks for experimentation, forecasting, and performance measurement.
- Build and maintain data infrastructure that ensures data quality, consistency, and accessibility for analytical workflows.
- Collaborate with product engineering, product, and operations teams to translate business requirements into production-grade data systems and insights.
- Conduct A/B tests, causal analyses, and performance evaluations to drive measurable improvements in key metrics.
- Implement monitoring, alerting, and automation for data systems supporting real-time decision-making.
- Mentor team members on scalable data engineering and quantitative problem-solving best practices.
Requirements
- 4+ years of experience building production data pipelines and infrastructure at scale.
- Strong proficiency in Python, SQL, and distributed computing frameworks such as Spark, Flink, and Hadoop.
- Expertise in statistical methods, predictive modeling, hypothesis testing, and experimental design.
- Solid understanding of cloud services for data storage, processing, and orchestration.
- Bachelor's or master's degree in Computer Science, Statistics, Applied Mathematics, or a related quantitative field.
- Excellent problem-solving skills focused on delivering business impact through reliable systems.
Nice to Have
- Experience in consumer technology or social media.
- Experience with real-time streaming systems and low-latency data processing.
- Contributions to open-source data tools or publications on large-scale analytics systems.
- A track record of reducing operational costs or improving system efficiency through data optimizations.
- Ability to bridge engineering excellence with rigorous analytical approaches.
Compensation and Benefits
- Base salary: $180,000–$440,000 USD.
- Equity and comprehensive medical, vision, and dental coverage.
- 401(k) retirement plan.
- Short- and long-term disability insurance.
- Life insurance, discounts, and other perks.
Skills
Python, SQL, Spark, Apache Kafka, Apache Flink, Hadoop, Statistical Modeling, Predictive Modeling, A/B Testing, Causal Analysis, Cloud Computing, Data Orchestration, Real-Time Streaming, Monitoring
Similar jobs
Data Engineering jobsOwn the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.
Builds and optimizes scalable data pipelines, storage, and OLAP databases for ML training, analytics, and product features. Requires 5+ years in data engineering, proficiency in Python/SQL/cloud platforms, and distributed systems experience.
Build and operate scalable data infrastructure, including partner data sharing, identity graph foundations, and governed batch and real-time platforms. The role requires 5+ years of data, distributed systems, infrastructure, or backend engineering experience and strong cloud and data-platform expertise.
Builds scalable data pipelines and data engine architecture for machine learning, integrating foundation models to automate labeling and discovery. The role requires 5+ years of experience, modern ML infrastructure expertise, and U.S. citizenship with security-clearance eligibility.
Build and operate reliable, production-grade data pipelines, warehouse infrastructure, and trusted datasets supporting company-wide analytics and AI initiatives. The role requires 3+ years of production data engineering experience, strong SQL and Python skills, and experience with Snowflake, dbt, cloud infrastructure, and orchestration.