Skip to content
PinterestPinterest

Engineering Manager II, Big Data Storage

Staff-level engineer leading design and development of Pinterest’s exabyte-scale data lake storage platform using Iceberg and related big data technologies to support ML/AI workloads.

About the job

What you’ll do:

  • Design, implement, and optimize Pinterest’s exabyte-scale data lake storage platform.
  • Lead complex technical projects and initiatives for data lake storage and metadata management, driving them from architecture through execution.
  • Collaborate with stakeholders and partner teams across the organization to design storage and metadata layer technologies that unlock big data and ML/AI innovations.
  • Build storage capabilities that efficiently support large-scale ML/AI workloads, including high-throughput data access, schema evolution, and large-scale column backfills.
  • Shape the long-term technical direction for scalable, reliable, and efficient big data storage systems.
  • Engage with and contribute to open source communities such as Apache Iceberg, Spark, and Flink to help address Pinterest’s scaling challenges.

What we’re looking for:

  • 8+ years of relevant industry experience designing and building large-scale production distributed systems.
  • Strong experience designing and maintaining scalable storage, metadata, or data lake infrastructure.
  • Experience building storage capabilities for large-scale ML/AI or analytics workloads, including high-throughput data access, schema evolution, and large-scale column backfills.
  • Deep knowledge with building distributed systems, data storage systems, and production infrastructure.
  • Experience with big data technologies such as Apache Iceberg, Spark, Flink, Presto/Trino, Hive, or similar systems.
  • Proficiency in programming languages like Java, Scala, or Python.
  • Proven ability to lead complex technical initiatives and influence architecture across teams.
  • Strong collaboration, communication, and problem-solving skills, with a drive for technical excellence and innovation.
  • Bachelor’s degree in a relevant field such as Computer Science, or equivalent experience

Skills

Apache Iceberg, Spark, Apache Flink, Presto, Trino, Apache Hive, Java, Scala, Python, Distributed Systems

Fetch

Fetch

United States

Senior Software Engineer II, Data Platform
$176k+/yrRemote8+ YOEData Engineering

Senior software engineer building and evolving Fetch’s data platform, including pipelines, governed data access, delivery infrastructure, and partner integrations. The role requires 8+ years of experience, strong platform or backend expertise, ownership of complex cross-team initiatives, and excellent technical judgment.

Thyme Care

Thyme Care

Remote

Senior Platform Engineer, Data
$176k+/yrRemote5+ YOEData Engineering

The Senior Platform Engineer will build and operate reliable data platform tooling, consolidate orchestration, scale dbt infrastructure, and improve Databricks developer experience. The role requires 5+ years of production software experience, strong Python and AWS expertise, infrastructure-as-code experience, and familiarity with modern data stacks.

Runpod

Runpod

San Francisco, CA

Senior Data Engineer
$175k+/yrRemote5+ YOEData Engineering

The Senior Data Engineer will design scalable data pipelines and warehousing systems supporting analytics, business metrics, and machine-learning initiatives. The role requires 4+ years of enterprise data experience, expertise with modern data platforms and ETL, and the ability to mentor engineers and collaborate across functions.

Creditgenie

Creditgenie

Plymouth Meeting, PA
Senior Data Engineer
$180k+/yrOn-site5+ YOEData Engineering

Senior Data Engineer responsible for designing and operating scalable data pipelines and platform capabilities across Snowflake and AWS. The role requires 5+ years of production data engineering experience, strong SQL and Python skills, and expertise in ETL/ELT, orchestration, quality, and observability.

Machinify

Machinify

United States

Senior Data Engineer 
$180k+/yrRemote6+ YOEData Engineering

Build and scale production data pipelines that transform raw healthcare and customer data into trusted canonical datasets powering ML models, dashboards, and product decisions. The role requires 6+ years of data engineering experience and strong Python, Spark SQL, and Airflow expertise.