Staff Data Engineer
Staff Data Engineer building and scaling data pipelines, integrations, and workflow orchestration systems. Owns architecture, IaC strategy, and technical leadership across large-scale data infrastructure.
About the job
What You'll Do
- Define the architecture and long-term technical direction, design and implementation of our data pipelines and integrations platform for storage, transformation and export at scale.
- Drive integrations with third-party tools and establish the standards that make future integrations faster and more reliable.
- Own our infrastructure as code strategy using Terraform, and contribute to how we evolve our cloud data infrastructure.
- Architect workflow orchestration systems for near-real-time and batch data processing, with a focus on reliability and scalability.
- Lead the design and implementation of data export pipelines that serve diverse customer needs.
- Spearhead development of internal tooling and agentic workflows that meaningfully accelerate engineering velocity across the org.
- Serve as a technical anchor — leading design reviews, mentoring engineers, and elevating how the team approaches complex problems.
- Partner with engineering leadership on roadmap prioritization, cross-team dependencies, and org-wide technical strategy.
- Communicate clearly about technical decisions, tradeoffs, and project status to both engineers and non-technical stakeholders.
About You
- Deep expertise building and scaling data pipelines, and a track record of owning complex data infrastructure end-to-end.
- AI-driven developer who dives deep into AI first workflows.
- Expert Python engineer with strong opinions about how to build reliable, maintainable systems at scale.
- Designed and led large-scale third-party API integration work, and you know what separates a maintainable integration from a brittle one.
- Fluent in infrastructure as code (Terraform, CloudFormation, or similar) and think about infrastructure as a product, not an afterthought.
- Experienced with AI/LLM tool integrations (Claude, Copilot, etc.) and understand the unique infrastructure demands they create.
- Production experience with workflow orchestration tools (Prefect, Airflow, Dagster) and can make the hard architectural calls.
- Natural technical leader — you build consensus, unblock others, and make the engineers around you better.
- Think in systems: you consider user needs, business impact, and long-term maintainability when designing solutions.
- Exceptional communicator who can write a design doc, lead a review, and distill a complex tradeoff for any audience.
- Located in EST or CST time zone.
Bonus Points
- Thrived at a rapidly scaling startup and know how to balance speed with technical rigor.
- Hands-on experience with Delta Lake and lakehouse architectures.
- Built deep integrations with developer tools (Jira, GitHub, GitLab, CI/CD systems) and have strong opinions about how to do it well.
- Strong perspectives on how engineering teams work best — and the tools that help or hurt.
Skills
Python, Terraform, CloudFormation, Prefect, Airflow, Dagster, Delta Lake, Ai/Llm Integrations, Infrastructure As Code, Workflow Orchestration
Similar jobs
Data Engineering jobsStaff Software Engineer leading design and development of large-scale batch and real-time data pipelines and ML infrastructure to power GenAI/LLM products and features for Airbnb's Messaging, Notifications, and Connectivity organization. Requires 9+ years experience building production ML systems and cross-functional collaboration.
This staff-level data engineer will architect and operate low-latency market data infrastructure, including feed handling, normalization, distribution, and exchange connectivity. The role requires at least five years of backend engineering experience and strong Java or C++ expertise with high-throughput messaging and market data protocols.
Leads large-scale advertising data ingestion, measurement, and agentic workflow systems, combining deep ad-tech expertise with production LLM experience. Requires 10+ years of engineering experience and technical and people leadership in complex enterprise environments.
Staff Software Engineer responsible for architecting, building, and operating Commure’s data warehouse platform, including CDC, lakehouse, query, transformation, and analytics layers. Requires 6+ years of software engineering experience and broad expertise across modern production data infrastructure.
Senior individual contributor responsible for architecting and scaling production data ingestion systems that integrate complex enterprise sources into reliable datasets. Requires 5+ years of backend engineering experience, strong Python, cloud, Kubernetes, Postgres, and data integration expertise.