Latest Data Engineering jobs
Job results
Leads the re-platforming of Vanta’s compliance data layer from MongoDB to schema-aware PostgreSQL across high-throughput Kafka and S3 pipelines. The role requires staff-level distributed systems expertise, migration leadership, and strong experience with relational and document data modeling.
Build and scale AWS-based data infrastructure, pipelines, knowledge graphs, and APIs for a CTV performance advertising platform. The role requires production data engineering experience with Spark, Scala, AWS, SQL, and large-scale services, plus a bachelor's degree.
Leads the strategy, architecture, and hands-on development of scalable AWS data ingestion and transformation platforms. Requires expert Python and SQL skills, Terraform and cloud-native pipeline experience, and 8+ years in data engineering or backend development, with technical leadership responsibilities.
Leads the operational engine for collecting and annotating real-world and simulated data used by perception and robot-learning teams. The role manages vendors and annotators, quality systems, dataset governance, dashboards, and cross-functional delivery.
Build and own production data pipelines, knowledge graph data models, and structured datasets from messy sources (PDFs, spreadsheets, telemetry) to power internal tools, dashboards, and ML models at a frontier AI compute infrastructure company. Requires experience operating depended-on pipelines, schema modeling, data quality engineering, and unstructured data extraction.
The Senior Analytics Engineer will architect and operate marketing data infrastructure, productionize predictive models, and enable attribution, experimentation, and customer activation. The role requires strong Snowflake, dbt, Python, SQL, and Segment expertise, plus experience with marketing data and cross-functional analytics initiatives.
Own the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.
Supports data platform improvements, modeling, analysis, and pipeline maintenance for corporate finance and analytics. The internship suits a quantitative bachelor’s or master’s student with SQL, database, and data visualization knowledge, with Python and cloud warehouse experience valued.
Build and operate distributed data applications powering large-scale audience segmentation and real-time personalization. The role requires 2–4 years of software engineering experience, backend development skills, and familiarity with databases, algorithms, and distributed systems.
Build and operate large-scale data acquisition pipelines, distributed processing systems, and production Java services across batch and streaming workloads. The role requires 5+ years of backend or data engineering experience, strong distributed-systems expertise, and proficiency with cloud data technologies.
Build and operate scalable monetization data platforms, pipelines, models, and quality systems spanning product, financial, and operational data. The role partners with Product Engineering, Finance, Accounting, Analytics, and GTM teams to deliver reliable, observable data products.
Leads an analytics engineering team that transforms raw data into reliable, actionable insights for product, marketing, and operations. The role requires 7+ years in data or analytics engineering, management experience, and advanced SQL, Databricks, and dbt expertise.
This staff-level data engineer will architect and operate low-latency market data infrastructure, including feed handling, normalization, distribution, and exchange connectivity. The role requires at least five years of backend engineering experience and strong Java or C++ expertise with high-throughput messaging and market data protocols.
Leads the analytics engineering function, owning data architecture, modeling standards, semantic layers, governance, and roadmap execution while managing and developing the team. The role requires deep SQL and dbt expertise, dimensional modeling experience, cloud data warehouse knowledge, and strong senior-stakeholder communication.
Architects scalable data systems and platforms using distributed technologies like Spark, Kafka, and AWS. Mentors engineers and drives innovation on large-scale data projects, requiring 8+ years experience and expertise in data infrastructure.
Build and optimize Ray Data, a Python-native data processing engine for large-scale AI workloads. The role focuses on distributed systems performance, scalable data pipelines, production training solutions, and fault tolerance while partnering with AI-focused customers.
Own and evolve Bevi’s end-to-end data platform, from ingestion and IoT modeling through governed self-service analytics and AI access. The senior individual contributor will architect scalable streaming and batch systems, establish governance and observability, and provide technical leadership across the Data & Data Science organization.
Builds agentic AI, automated data workflows, and BI solutions for complex telecommunications datasets. The role requires 5+ years of technical data and automation experience, strong SQL and Python skills, and expertise in data governance and LLM-based tools.
The Senior Data and AI Specialist will build agentic AI solutions, automated Python workflows, and analytics products across telecommunications data. The role requires at least five years of experience with SQL, data automation, BI tools, LLM agents, and data governance.
Leads the design and delivery of scalable data infrastructure, services, and developer tooling while shaping technical direction and mentoring engineers. Requires 8+ years of software engineering experience, strong systems design expertise, and proficiency in a modern programming language.
Senior Software Engineer on Observe by Snowflake’s Data Management team, owning the APIs, schemas, and abstractions for scalable tables, views, and materialized views across streaming telemetry. Requires 5+ years of experience with databases, SQL, streaming or data pipelines, API design, and production distributed systems.
Owns marketing-sourced pipeline across multiple growth motions by designing, launching, and optimizing integrated B2B SaaS campaigns. The role requires 10+ years in demand generation or growth marketing, pipeline ownership, strong funnel analytics, and hands-on experience with Salesforce, Marketo, and ABM platforms.
Builds customer intelligence workflows that turn product usage, engagement, and contract data into actionable insights for Customer Success. The role requires strong Python and SQL skills, a bachelor's degree, and an interest in applied AI, automation, and predictive customer health modeling.
Senior software engineer building and evolving Fetch’s data platform, including pipelines, governed data access, delivery infrastructure, and partner integrations. The role requires 8+ years of experience, strong platform or backend expertise, ownership of complex cross-team initiatives, and excellent technical judgment.
Builds the data foundation for a client retention and personalization team, including ELT pipelines, predictive data models, marketing integrations, and measurement layers. Requires modern data-stack experience, strong SQL and programming skills, and regular use of AI coding assistants.
Build and own Wrapbook’s analytics layer, including production pipelines, governed data models, canonical datasets, self-serve analytics, and monitoring. The role requires strong SQL and Python, modern warehouse experience, and 4+ years in data or analytics engineering.
Staff Software Engineer leading design and development of large-scale batch and real-time data pipelines and ML infrastructure to power GenAI/LLM products and features for Airbnb's Messaging, Notifications, and Connectivity organization. Requires 9+ years experience building production ML systems and cross-functional collaboration.
Build and maintain data pipelines, analytics models, dashboards, and external data products while partnering with engineering, product, implementation teams, and customers. The role requires 5+ years of analytics or data engineering experience, strong dbt and SQL expertise, and customer-facing collaboration skills.
The Senior Software Engineer II will build and operate integrations, APIs, data pipelines, and data contracts connecting Wrapbook with accounting, ERP, data, and AI ecosystems. The role requires strong data architecture and coding experience with Python, Rails, and modern integration tooling.
Designs and builds scalable distributed data systems and pipelines powering blockchain analytics products. The role is hands-on, requiring production experience with Scala, Java, or Python, big-data technologies, cloud infrastructure, and data orchestration, plus technical leadership and mentoring.
Builds the account, contact, hierarchy, enrichment, and identity systems that power go-to-market operations. The role requires modern data-stack experience, production LLM development, entity resolution, SaaS integrations, and close partnership with Sales, Marketing, and RevOps.
Build and own the unified data layer powering internal analytics, customer-facing dashboards, benchmarks, and future product intelligence. The role requires production data-pipeline experience, strong SQL and Python, data modeling expertise, and ownership of architecture and data quality.
Leads the analytics engineering organization and sets strategy for Webflow’s enterprise data architecture, semantic layer, governance, and AI-native data systems. Requires 10+ years of data and analytics engineering experience, strong modern data stack expertise, and demonstrated team leadership.
Builds and operates large-scale web crawling, extraction, and data engineering infrastructure processing billions of pages. The role requires 5+ years of software engineering experience, strong distributed-systems fundamentals, and proficiency with Java or Python, cloud platforms, Kubernetes, and ETL technologies.
Leads technical direction for a research data platform, building scalable pipelines, platform components, and canonical datasets used by ML researchers. The role requires experience with data-intensive systems, schema design, cross-team technical leadership, and hands-on software development.
Builds and maintains scalable data platforms and pipelines that support analytics and AI/ML systems. The role requires 5+ years of data engineering or full-stack experience, strong SQL and ETL expertise, cloud infrastructure experience, and a bachelor's degree.
Build and operate the data foundations behind AI safeguards, including production pipelines, data stores, governance controls, and internal tooling. The role requires strong Python and SQL skills, production data-platform experience, and expertise in reliable, privacy-conscious systems across multiple cloud environments.
Build customer-facing data products and shared platform systems that transform conflicting, constantly changing sources into reliable, searchable information. The role requires 8+ years of hands-on engineering experience, strong Python and SQL skills, and ownership of product quality, reliability, and delivery.
Manages the full lifecycle of scientific research data, including governance, metadata, repositories, open-science publishing, and AI/ML compatibility. The role also builds partnerships with federal research organizations and requires a bachelor’s degree plus 5–7+ years of relevant experience.
The Senior Data Engineer will build scalable pipelines, ETL workflows, and data products while partnering with analytics, product, and engineering teams. The role requires 8–10+ years of experience, strong SQL and programming skills, cloud data-platform expertise, and a focus on reliability, quality, and automation.
The Senior Platform Engineer will build and operate reliable data platform tooling, consolidate orchestration, scale dbt infrastructure, and improve Databricks developer experience. The role requires 5+ years of production software experience, strong Python and AWS expertise, infrastructure-as-code experience, and familiarity with modern data stacks.
Build production-grade data processing systems and performant datasets that support AI-driven customer experiences and company-wide analytics. The role requires 5+ years of data engineering experience, strong Python and SQL skills, and expertise with distributed processing frameworks and modern analytics tooling.
The Senior Data Engineer will design scalable data pipelines and warehousing systems supporting analytics, business metrics, and machine-learning initiatives. The role requires 4+ years of enterprise data experience, expertise with modern data platforms and ETL, and the ability to mentor engineers and collaborate across functions.
Builds scalable data pipelines and data engine architecture for machine learning, integrating foundation models to automate labeling and discovery. The role requires 5+ years of experience, modern ML infrastructure expertise, and U.S. citizenship with security-clearance eligibility.
The Senior Database Administrator will design, operate, secure, and optimize high-volume PostgreSQL databases supporting a scalable SaaS platform. The role requires 6+ years managing databases at scale, strong performance-tuning expertise, and experience with high availability, automation, observability, and recovery.
Build and own foundational data models, reliable pipelines, and data platforms that support analytics and business decision-making. The role requires 5+ years of data engineering experience, strong Python and SQL skills, cloud data warehouse expertise, and a bachelor's degree or equivalent experience.
Staff engineer owning the analytical data layer, schema, and tiered analytics architecture. The role combines hands-on backend development with database performance optimization, observability, ingestion coordination, and measured architectural decision-making.
Builds and operates reliable, scalable data pipelines and infrastructure for Auth0’s data platform. The role requires 5+ years of software development experience, strong SQL and Python skills, and hands-on experience with cloud data systems and modern data-stack tools.
Leads Alpaca’s data department across platform engineering, analytics engineering, and data science, owning strategy, architecture, execution, and operational reliability. The role requires extensive data engineering and people-management experience, modern data-stack expertise, and familiarity with financial services data.
Build and operate petabyte-scale data infrastructure powering Discord’s insights and products. The role requires 5+ years of software engineering experience, strong programming skills, and experience with large-scale pipelines, streaming, orchestration, or data warehousing.