Senior Data Engineer - Revenue Data Platform
Build and operate large-scale revenue data pipelines powering billing and cost attribution, while improving reliability, latency, and correctness. The role requires strong Spark and Airflow experience, cross-functional problem-solving, and operational ownership of mission-critical production systems.
About the job
Responsibilities
- Design and build high-throughput data pipelines for billing and cost attribution.
- Drive platform improvements, including latency reduction, Spark optimization, sharding, and cross-datacenter reliability.
- Own root-cause investigations on billing accuracy issues in collaboration with Finance and Product teams.
- Contribute to new billing features.
- Participate in an on-call rotation and maintain a high reliability bar for production systems.
- Contribute to engineering standards and help grow the technical culture of the team.
Requirements
- Significant experience building and operating production data pipelines at scale using Spark and Airflow.
- Experience owning complex, cross-functional investigations and driving them to resolution.
- Ability to balance feature delivery with platform health and technical debt.
- Strong analytical instincts and rigor around correctness in data systems.
- Ability to work effectively in ambiguous environments and drive alignment across stakeholder teams.
- Commitment to code quality, maintainability, and operational excellence.
Nice-to-haves
- Experience with Apache Iceberg or lakehouse architectures.
- Experience with billing, metering, or financial data systems.
- Experience with low-latency data serving or real-time aggregation pipelines.
Compensation and Benefits
- Estimated yearly salary: $192,000–$240,000 USD.
- Competitive salary and equity package; variable compensation may be included.
- Healthcare, dental, parental planning, and mental health benefits.
- 401(k) plan and match, paid time off, fitness reimbursements, and discounted employee stock purchase plan.
Skills
Python, Scala, Spark, Apache Airflow, Trino, Apache Iceberg, Data Pipelines, Sharding, Lakehouse Architecture, Real-Time Aggregation
Similar jobs
Data Engineering jobsBuild and operate low-latency systems that capture, normalize, and distribute real-time market data for institutional trading. The role requires backend engineering experience, Java or C++, market data infrastructure knowledge, and exchange connectivity expertise.
Build and own large-scale data models, batch and real-time pipelines, and data infrastructure that provide reliable datasets and insights across Plaid. The role requires 4+ years of data engineering experience, strong SQL and Python skills, and expertise with modern warehouses, lakes, and orchestration tools.
Builds and owns scalable SQL/Python data pipelines, golden datasets, and workflows using DBT, Airflow, Redshift for large-scale data (500TB+). Collaborates cross-functionally to enable data-driven decisions at Plaid. Requires 4+ years data engineering experience.
Build and lead the central data platform, covering ingestion, warehousing, orchestration, streaming, self-service frameworks, and trust layers. The role requires 5+ years of production data infrastructure experience, strong Python and SQL skills, and expertise with Snowflake and modern data tooling.
Build and operate Jump’s Snowflake and dbt data platform, including ELT pipelines, governance, testing, monitoring, and analytics models. The role requires 6+ years of data engineering experience, production-grade SQL and Snowflake expertise, deep dbt knowledge, and strong Python and software engineering practices.