Staff Software Engineer, Data Catalog
Staff Software Engineer leading development of data catalog and metadata infrastructure for discovery, governance, lineage, and quality across Airbnb’s data ecosystem. Requires 9+ years of software engineering experience focused on data infrastructure and strong programming and distributed data technology skills.
About the job
Responsibilities
- Develop and implement data catalog products for data discovery, metadata management, data quality, and data lineage.
- Build metadata infrastructure supporting data governance, including ownership management, data classification, privacy, access control, and retention.
- Build an end-to-end data lineage product spanning online, offline, machine learning, metrics, visualizations, and third-party datasets.
- Collaborate with cross-functional teams to catalog and govern datasets and meet business and regulatory requirements.
- Define metadata quality and governance requirements and develop processes to enforce them.
- Design and build metadata integrations across data infrastructure systems.
- Build scalable metadata onboarding tools and integration processes.
- Implement metadata-driven data policies and procedures.
Requirements
- Bachelor’s, master’s, or doctoral degree in Computer Science or a related field, or equivalent experience.
- 9+ years of software engineering experience focused on data infrastructure.
- Experience with data storage and distributed processing technologies such as Hive, Spark, Trino, Flink, or SQL databases.
- Strong programming skills in Java, Python, or Scala.
- Experience with data catalog products, metadata infrastructure, metadata management, metadata integration frameworks, and metadata-driven data management.
- Experience with workflow orchestration solutions such as Apache Airflow, Prefect, or Kubeflow.
- Strong knowledge of data governance frameworks, including data classification, lineage, quality, privacy, and retention.
- Excellent communication, collaboration, analytical, and problem-solving skills.
Compensation
- Base salary range: $212,000–$265,000 USD annually.
- May also be eligible for bonus, equity, benefits, and employee travel credits.
Skills
Data Catalog, Metadata Management, Data Lineage, Data Governance, Data Quality, Spark, Apache Hive, Trino, Apache Flink, SQL, Java, Python, Scala, Apache Airflow, Prefect
Similar jobs
Data Engineering jobsStaff Software Engineer responsible for architecting, building, and operating Commure’s data warehouse platform, including CDC, lakehouse, query, transformation, and analytics layers. Requires 6+ years of software engineering experience and broad expertise across modern production data infrastructure.
Senior individual contributor responsible for architecting and scaling production data ingestion systems that integrate complex enterprise sources into reliable datasets. Requires 5+ years of backend engineering experience, strong Python, cloud, Kubernetes, Postgres, and data integration expertise.
This staff-level data engineer will architect and operate low-latency market data infrastructure, including feed handling, normalization, distribution, and exchange connectivity. The role requires at least five years of backend engineering experience and strong Java or C++ expertise with high-throughput messaging and market data protocols.
Analytics Engineer supporting Go-to-Market teams by building scalable data models, metrics, pipelines, visualizations, and self-service products. The role requires 10+ years of data experience, deep SQL expertise, Python proficiency, and strong business judgment.
Staff Software Engineer leading design and development of large-scale batch and real-time data pipelines and ML infrastructure to power GenAI/LLM products and features for Airbnb's Messaging, Notifications, and Connectivity organization. Requires 9+ years experience building production ML systems and cross-functional collaboration.