Staff Software Engineer, Business Data
Leads technical delivery for scalable data systems, pipelines, warehouses, and supporting applications while mentoring engineers and partnering cross-functionally. Requires extensive experience with distributed data frameworks, backend development, SQL, data quality, and technical leadership.
About the job
Responsibilities
- Lead the technical outcomes for a team of engineers, providing mentorship, guidance, and support.
- Partner with recruiting to attract and hire top talent.
- Deliver reliable, efficient data pipelines that scale to user needs.
- Develop subject-matter expertise and manage SLAs for data pipelines and full-stack web applications supporting critical stakeholders.
- Collaborate with product managers and peers to create and improve canonical datasets and data warehouses, use golden paths, and ensure trustworthy data.
- Leverage AI, LLMs, and agents at scale to produce and analyze high-quality data on ambiguous problems.
- Drive key data initiatives through the full development lifecycle, from planning through delivery, while maintaining high quality and timely completion.
- Foster a collaborative, inclusive environment that promotes innovation, knowledge sharing, and continuous improvement.
Requirements
- Typically 10+ years of experience building and operating data systems, pipelines, warehouses, infrastructure, and leading teams.
- Strong engineering background and passion for data, with experience writing and debugging pipelines using distributed data frameworks.
- Ability to investigate data inconsistencies, identify root causes, and resolve data-quality issues.
- Knowledge of a backend development language such as Scala, Java, or Go.
- Strong SQL experience.
- Customer-focused approach and ability to partner with product leaders, business stakeholders, and engineers.
- Effective cross-functional collaboration, rigorous thinking, clear communication, and sound decision-making.
- Ability to thrive with autonomy and responsibility in an ambiguous environment.
- Ability to foster and contribute to a healthy, inclusive, challenging, and supportive work environment.
Nice-to-haves
- Experience with Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, or AWS Cloud.
- Contributions to open-source projects.
- Experience creating and maintaining data marts or warehouses for business reporting.
- Experience collaborating with Product, Go-To-Market, Sales, or Marketing teams.
- Interest in innovation and understanding how systems work, with the ability to question and direct architectural decisions.
- Strong written and verbal communication skills for leadership, users, and company-wide audiences.
Skills
Apache Airflow, Spark, Apache Kafka, Apache Flink, Scala, Java, Go, SQL, Apache Iceberg, Change Data Capture, AWS, Apache Trino, Apache Pinot, Data Warehousing, LLMs
Similar jobs
Data Engineering jobsLeads enterprise data engineering strategy, architecture, delivery, governance, and technical leadership across the organization. Requires extensive data engineering experience, advanced data modeling and warehouse expertise, and strong PySpark, SQL, and Python skills.
Staff Software Engineer leading development of data catalog and metadata infrastructure for discovery, governance, lineage, and quality across Airbnb’s data ecosystem. Requires 9+ years of software engineering experience focused on data infrastructure and strong programming and distributed data technology skills.
Staff Software Engineer responsible for designing and scaling Ray Data’s distributed data-processing infrastructure for large-scale AI training and inference. Requires 6+ years of production software and architectural ownership experience, plus deep distributed-systems expertise and strong Python skills.
Leads the design, operation, and technical direction of Pinterest’s data workflow and context control planes, driving reliability, scalability, AI-native capabilities, and open-source contributions. Requires 10+ years of distributed-systems experience, infrastructure expertise, and proficiency in Python or Java.
Leads the design and improvement of large-scale data platform systems and workflows, collaborating across engineering, data science, and business teams. Requires 10+ years of relevant experience, advanced SQL and Python, cloud data tooling, and technical leadership.