Staff Software Engineer
Staff Software Engineer building and scaling Plaid's Data Infrastructure platform (warehouses, lakehouses, Spark, streaming, orchestration). Lead projects to improve ML workflows, data freshness, ETL pipelines; mentor engineers and reduce operational burden. Requires 6+ years software engineering with deep data infrastructure expertise.
About the job
Responsibilities
- Contribute towards the long-term technical roadmap for data-driven and machine learning iteration at Plaid.
- Lead key data infrastructure projects such as improving ML development golden paths, implementing offline streaming solutions for data freshness, building net new ETL pipeline infrastructure, and evolving data warehouse or data lakehouse capabilities.
- Work with stakeholders in other teams and functions to define technical roadmaps for key backend systems and abstractions across Plaid.
- Debug, troubleshoot, and reduce operational burden for our Data Platform.
- Grow the team via mentorship and leadership, reviewing technical documents and code changes.
Requirements
- 6+ years of software engineering experience.
- Extensive hands-on software engineering experience, with a strong track record of delivering successful projects within the Data Infrastructure or Platform domain at similar or larger companies.
- Deep understanding of one of the below: Data Infrastructure systems, including Data Warehouses, Data Lakehouses, Apache Spark, Streaming Infrastructure, Workflow Orchestration.
- Strong cross-functional collaboration, communication, and project management skills, with proven ability to coordinate effectively.
- Proficiency in coding, testing, and system design, ensuring reliable and scalable solutions.
- Demonstrated leadership abilities, including experience mentoring and guiding junior engineers.
Nice-to-Haves
- Experience with Databricks.
- Experience with Airflow.
- Experience with AWS EMR.
- Experience with Python.
Skills
Data Warehouses, Data Lakehouses, Spark, Streaming Infrastructure, Workflow Orchestration, Databricks, Airflow, Aws Emr, Python, ETL, Machine Learning
Similar jobs
Data Engineering jobsThis staff-level data engineer will architect and operate low-latency market data infrastructure, including feed handling, normalization, distribution, and exchange connectivity. The role requires at least five years of backend engineering experience and strong Java or C++ expertise with high-throughput messaging and market data protocols.
Staff Software Engineer responsible for architecting, building, and operating Commure’s data warehouse platform, including CDC, lakehouse, query, transformation, and analytics layers. Requires 6+ years of software engineering experience and broad expertise across modern production data infrastructure.
Senior individual contributor responsible for architecting and scaling production data ingestion systems that integrate complex enterprise sources into reliable datasets. Requires 5+ years of backend engineering experience, strong Python, cloud, Kubernetes, Postgres, and data integration expertise.
Staff Software Engineer leading design and development of large-scale batch and real-time data pipelines and ML infrastructure to power GenAI/LLM products and features for Airbnb's Messaging, Notifications, and Connectivity organization. Requires 9+ years experience building production ML systems and cross-functional collaboration.
Staff Software Engineer leading development of data catalog and metadata infrastructure for discovery, governance, lineage, and quality across Airbnb’s data ecosystem. Requires 9+ years of software engineering experience focused on data infrastructure and strong programming and distributed data technology skills.