Skip to content
StripeStripe

Staff Software Engineer, Data Quality and Governance

Leads engineers building and operating large-scale data discovery, metadata, catalog, pipeline, and warehouse systems while improving data quality and governance. Requires 10+ years of data-systems experience, distributed-systems expertise, backend programming, strong SQL, and technical leadership.

About the job

Responsibilities

  • Lead technical outcomes for a team of engineers through mentorship, guidance, and support.
  • Build and operate large-scale data discovery, metadata, and catalog platforms.
  • Develop subject-matter expertise and manage SLAs for data pipelines and full-stack web applications supporting critical stakeholders.
  • Collaborate with product managers and cross-functional peers to create and improve canonical datasets and data warehouses, establish golden paths, and ensure trustworthy data.
  • Leverage AI, LLMs, and agents at scale to produce and analyze high-quality data for ambiguous problems.
  • Drive key data initiatives through the full development lifecycle, from planning to delivery, while maintaining high standards of quality and timely completion.
  • Foster a collaborative, inclusive, innovative, and supportive engineering environment.

Requirements

  • 10+ years of experience building and operating data systems, pipelines, warehouses, or infrastructure, and leading teams to deliver solutions.
  • Strong distributed-systems fundamentals.
  • Ability to investigate data inconsistencies, identify root causes, and resolve data-quality issues.
  • Proficiency in a backend development language such as Scala, Java, or Go.
  • Strong SQL experience.
  • Customer-focused approach and ability to partner with product leaders, business stakeholders, and engineers.
  • Strong cross-functional collaboration, communication, judgment, and decision-making skills.
  • Ability to work autonomously and effectively in an ambiguous environment.
  • Commitment to fostering a healthy, inclusive, challenging, and supportive workplace.

Nice-to-haves

  • Experience with Iceberg, Kafka, change data capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, or AWS Cloud.
  • Experience influencing open-source contributions.
  • Experience creating and maintaining data marts or warehouses for business reporting.
  • Experience collaborating with Product, Go-To-Market, Sales, or Marketing teams.
  • Strong interest in innovation and architectural decision-making.
  • Strong written and verbal communication skills for technical, leadership, user, and company-wide audiences.

Skills

Distributed Systems, SQL, Scala, Java, Go, Apache Iceberg, Apache Kafka, Change Data Capture, Apache Flink, Spark, Apache Airflow, Hive Metastore, Pinot, Trino, Aws Cloud

SmithRx

SmithRx

United States

Senior Staff Data Engineer
$179k+/yrRemote12+ YOEData Engineering

Leads enterprise data engineering strategy, architecture, delivery, governance, and technical leadership across the organization. Requires extensive data engineering experience, advanced data modeling and warehouse expertise, and strong PySpark, SQL, and Python skills.

Airbnb

Airbnb

United States

Staff Software Engineer, Data Catalog
$212k+/yrRemote9+ YOEData Engineering

Staff Software Engineer leading development of data catalog and metadata infrastructure for discovery, governance, lineage, and quality across Airbnb’s data ecosystem. Requires 9+ years of software engineering experience focused on data infrastructure and strong programming and distributed data technology skills.

Anyscale

Anyscale

San Francisco, CA

Staff Software Engineer, Ray Data
$240k+/yrOn-site7+ YOEData Engineering

Staff Software Engineer responsible for designing and scaling Ray Data’s distributed data-processing infrastructure for large-scale AI training and inference. Requires 6+ years of production software and architectural ownership experience, plus deep distributed-systems expertise and strong Python skills.

Pinterest

Pinterest

United States

Staff Software Engineer, Workflow Platform
$177k+/yrRemote10+ YOEData Engineering

Leads the design, operation, and technical direction of Pinterest’s data workflow and context control planes, driving reliability, scalability, AI-native capabilities, and open-source contributions. Requires 10+ years of distributed-systems experience, infrastructure expertise, and proficiency in Python or Java.

Mozilla

Mozilla

United States
Senior Staff Data Engineer
No salary listedHybrid10+ YOEData Engineering

Leads the design and improvement of large-scale data platform systems and workflows, collaborating across engineering, data science, and business teams. Requires 10+ years of relevant experience, advanced SQL and Python, cloud data tooling, and technical leadership.