Latest Data Engineering jobs
Job results
Leads enterprise data engineering strategy, architecture, delivery, governance, and technical leadership across the organization. Requires extensive data engineering experience, advanced data modeling and warehouse expertise, and strong PySpark, SQL, and Python skills.
Staff Software Engineer leading development of data catalog and metadata infrastructure for discovery, governance, lineage, and quality across Airbnb’s data ecosystem. Requires 9+ years of software engineering experience focused on data infrastructure and strong programming and distributed data technology skills.
Staff Software Engineer responsible for designing and scaling Ray Data’s distributed data-processing infrastructure for large-scale AI training and inference. Requires 6+ years of production software and architectural ownership experience, plus deep distributed-systems expertise and strong Python skills.
Build and maintain dbt models, Snowflake semantic layers, and ingestion pipelines across business functions while improving data quality and resilience. The role requires 4–6 years of analytics or data engineering experience, strong dbt and SQL expertise, and a quantitative bachelor's degree.
Leads the design, operation, and technical direction of Pinterest’s data workflow and context control planes, driving reliability, scalability, AI-native capabilities, and open-source contributions. Requires 10+ years of distributed-systems experience, infrastructure expertise, and proficiency in Python or Java.
Staff Software Engineer building scalable frameworks for high-performance financial data ingestion/distribution and AI-native products. Requires 7+ years experience with distributed systems, microservices, and data architectures; partners with product teams to drive technical direction.
Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.
Build scalable data ingestion, normalization, storage, and orchestration pipelines for multi-tenant device compliance data from endpoint-management platforms. The role requires 3+ years of data engineering experience, strong SQL and Python, database expertise, and experience with APIs and pipeline orchestration.
Build and govern quote-to-cash data models and products integrating Salesforce, CPQ, billing, and finance systems. The role requires 5+ years of data engineering experience, strong SQL and Python skills, and expertise in self-service analytics for GTM teams.
Analytics Engineering intern building dimensional data models, SQL pipelines, quality controls, and self-serve datasets or dashboards. Requires current quantitative-degree study, SQL proficiency, programming familiarity—preferably Python—and clear technical communication.
Data Engineering Intern supporting scalable pipelines and infrastructure for analytics and machine learning workloads. Requires Python and SQL proficiency, cloud familiarity, and exposure to modern software architecture or AI/API integrations.
Leads database architecture, performance, reliability, and developer-tooling initiatives for high-volume trading applications. Requires 8+ years of software engineering experience, expert MySQL skills, backend development expertise, and strong knowledge of distributed systems and database operations.
Build and operate foundational streaming, messaging, and data pipeline infrastructure for highly scalable identity and analytics systems. The role requires 3+ years of software development experience and strengths in distributed systems, event streaming, and platform reliability.
Senior Data Engineer responsible for building and operating reliable clinical and claims data pipelines, CDC systems, quality controls, and de-identified exports. The role requires 5+ years of production pipeline experience plus strong SQL, Python, Spark, and data-governance skills.
Leads the design and improvement of large-scale data platform systems and workflows, collaborating across engineering, data science, and business teams. Requires 10+ years of relevant experience, advanced SQL and Python, cloud data tooling, and technical leadership.
Own the company’s metric governance program by defining canonical metrics, enforcing them in semantic and catalog systems, improving data quality, and validating AI-agent outputs. Requires 5+ years in analytics or analytics engineering, strong SQL, production semantic-layer ownership, and experience with AI evaluation and data governance.
Construye y lidera la arquitectura, los pipelines y la plataforma de datos para habilitar analítica de producto, reportes financieros y experiencias self-serve. Requiere más de 7 años de experiencia, dominio de SQL, Snowflake, dbt, Python y AWS, además de experiencia con orquestación y modelado de datos.
Senior Data Engineer responsible for architecting and operating scalable data pipelines, warehouses, and analytics infrastructure. The role requires 7+ years of data or analytics engineering experience, strong SQL and modeling expertise, and proficiency with cloud, orchestration, and BI technologies.
Build and scale distributed data platforms, database systems, delivery services, and APIs, with emphasis on reliability, performance, observability, and data integrity. Requires 3+ years of software development experience with distributed systems and databases; Golang experience is preferred.
Own and evolve Stream’s revenue operations data platform, including ingestion, transformation, modeling, reliability, and GCP infrastructure. The role requires 6+ years of production data-platform experience, expert SQL, strong Python, modern ELT, Terraform, and technical leadership.
Build and operate distributed systems powering Apache Pinot’s real-time analytics platform at massive scale. The role requires strong distributed-systems expertise, end-to-end delivery ownership, and a focus on reliability, observability, and performance.
Own and scale transformation pipelines that convert diverse financial and operational data into reliable FP&A-ready models. The role requires strong SQL and dbt expertise, data integrity and performance skills, and effective collaboration across Engineering and Customer Success.
Oversee the lifecycle, quality, governance, and publication of research data across scientific programs. The role requires 3–5+ years of research data-management experience, strong metadata and FAIR-data expertise, and the ability to collaborate with researchers and engineers.
Staff Software Engineer responsible for architecting, building, and operating Commure’s data warehouse platform, including CDC, lakehouse, query, transformation, and analytics layers. Requires 6+ years of software engineering experience and broad expertise across modern production data infrastructure.
Build and operate low-latency systems that capture, normalize, and distribute real-time market data for institutional trading. The role requires backend engineering experience, Java or C++, market data infrastructure knowledge, and exchange connectivity expertise.
Senior Data Engineer responsible for designing and operating scalable data pipelines and platform capabilities across Snowflake and AWS. The role requires 5+ years of production data engineering experience, strong SQL and Python skills, and expertise in ETL/ELT, orchestration, quality, and observability.
Build and scale data pipelines, reusable datasets, and validation frameworks supporting business intelligence, marketing, and data science. The role requires strong Python and SQL skills, modern data-stack experience, and at least four years of software or data engineering experience.
Build and optimize scalable data pipelines, reusable datasets, and federated data quality systems for healthcare analytics. The role requires at least 2 years of data or software engineering experience and strong Python, SQL, AWS, orchestration, database, and warehouse expertise.
Leads the architecture, scaling, security, governance, and cost optimization of enterprise and AI data platforms. Requires 10+ years of data or software engineering experience, with expertise in production data foundations, CI/CD, governance, security, and performance optimization.
Build scalable, fault-tolerant data infrastructure and pipelines that transform web-scale corpora into training datasets for large language models. The role requires substantial distributed-systems experience, Apache Spark expertise, and strong Python or Rust skills.
Design and implement scientific data models, integrations, parsers, visualizations, and AI/ML-driven solutions for biopharma customers. The role combines customer-facing consulting in the Frankfurt region with product collaboration and requires deep life sciences expertise, German proficiency, and advanced industry experience.
Own and evolve trusted data models for Marketing and Product use cases, from design and testing through monitoring and documentation. The role requires 3–5 years of data or analytics engineering experience, strong SQL and Python, dbt expertise, and Snowflake or comparable warehouse experience.
Supports development of autonomous-vehicle safety risk models through large-scale driving and simulation data analysis, dataset pipelines, and empirical experimentation. Requires current enrollment in a quantitative B.S. or M.S. program, strong Python skills, and onsite availability for a part-time project.
Build and operate production data pipelines and transformation layers that turn heterogeneous business, identity, and fraud data into reliable inputs for entity resolution, scoring, and customer APIs. The role requires at least one year of data engineering experience with Python, SQL, cloud platforms, and modern pipeline tooling.
Own the architecture, reliability, and evolution of a modern data platform spanning data engineering and analytics/BI. The role requires 8+ years of experience plus deep expertise in SQL, dbt, ClickHouse, BigQuery, Looker, streaming systems, and distributed data platforms.
Build and own Stuut’s foundational data platform, including ingestion pipelines, canonical models, semantic layers, and observability. The role requires 3+ years of production data pipeline experience with Python, SQL, cloud warehouses, and ETL/ELT tooling.
Build and own large-scale data models, batch and real-time pipelines, and data infrastructure that provide reliable datasets and insights across Plaid. The role requires 4+ years of data engineering experience, strong SQL and Python skills, and expertise with modern warehouses, lakes, and orchestration tools.
Builds and owns scalable SQL/Python data pipelines, golden datasets, and workflows using DBT, Airflow, Redshift for large-scale data (500TB+). Collaborates cross-functionally to enable data-driven decisions at Plaid. Requires 4+ years data engineering experience.
Build and scale data ingestion platforms, pipelines, APIs, and processing products that move billions of rows across a multi-tenant system. The role requires 8+ years of software development experience and strong expertise in large-scale application architecture.
Architects and builds ZoomInfo’s distributed data-platform infrastructure, including federated GraphQL access, real-time pipelines, indexing, and observability. The role requires 10+ years of software engineering experience, cloud-native expertise, and strong distributed-systems design skills.
Leads the strategy, architecture, and scaling of Pinterest’s big data and AI infrastructure across petabyte-scale workloads. Requires principal-level technical leadership, extensive Kubernetes or big data platform experience, and proficiency in modern data and cloud technologies.
Senior individual contributor responsible for architecting shared dbt models, marketing attribution, data quality, and AI-driven analytics workflows. Requires 6+ years in analytics engineering or data, deep dbt and SQL expertise, modern data-stack experience, and strong marketing measurement knowledge.
Staff-level engineer responsible for the technical direction, reliability, and evolution of a cloud ELT platform supporting healthcare data products. The role requires 7+ years of software or data engineering experience, deep SQL/Python and modern data-platform expertise, and strong architectural and mentoring leadership.
Build scalable analytics engineering infrastructure, SaaS data models, and AI-enabled workflows that support enterprise decision-making. The role requires 3–6 years of hands-on analytics or data engineering experience, strong SQL and modern data modeling expertise, and cloud data warehouse experience.
Own the data foundation for machine-learning systems by building dataset pipelines, labeling workflows, quality controls, and lineage processes. The role requires at least three years of data engineering experience, strong Python and SQL skills, and familiarity with ML data quality concerns and orchestration tools.
Build simulation, performance-modeling, trade-study, and optimization tools that expand an AI-powered automated design engine for large-scale infrastructure projects. The role requires strong systems thinking, physics or engineering modeling experience, and hands-on software development; candidates at all experience levels are considered.
Builds and optimizes scalable data pipelines, storage, and OLAP databases for ML training, analytics, and product features. Requires 5+ years in data engineering, proficiency in Python/SQL/cloud platforms, and distributed systems experience.
Senior data platform engineer who scales infrastructure, automates data delivery, builds AI-enabled analytical tools, and leads cross-functional engineering initiatives. Requires 4+ years of data infrastructure experience, strong Kafka and distributed-systems expertise, and proficiency in Python, Scala, cloud platforms, and Terraform.
Senior data engineer who will set architecture standards, lead large-scale pipeline initiatives, and build reliable cloud-native data platforms on AWS. The role requires expertise in orchestration, batch and streaming systems, infrastructure as code, and modern data architectures.
Leads the data engineering team and platform strategy, overseeing pipelines, warehousing, governance, reliability, and data products supporting clinical operations and company decisions. Requires 7+ years across data engineering and people management, including 3+ years directly managing data engineers.