Skip to content

Staff Data Engineer

Builds and leads the development of large-scale distributed data systems and pipelines that power products and organizational decision-making. The role requires 7+ years of data engineering experience, cloud expertise, and strong technical leadership.

About the job

Responsibilities

  • Design, build, and support large-scale, distributed data systems that power products and services.
  • Develop data pipelines and optimize data storage and retrieval processes.
  • Ensure the reliability and scalability of data architecture.
  • Collaborate with cross-functional teams to understand data requirements and implement solutions.
  • Troubleshoot data-system issues.
  • Advocate for and implement software engineering best practices to improve the efficiency, maintainability, and robustness of data systems.
  • Provide technical leadership and apply problem-solving and analytical skills.

Requirements

  • BSc in Computer Science or a related discipline.
  • 7+ years of experience designing and building robust, scalable, distributed data systems and pipelines using open-source and public cloud technologies.
  • Strong experience with data orchestration tools such as Apache Airflow and Dagster.
  • Experience with big data storage and processing technologies such as dbt, Apache Spark, SQL, Athena/Trino, Amazon Redshift, Snowflake, and relational database management systems including PostgreSQL and MySQL.
  • Knowledge of event-driven architectures and streaming technologies such as Apache Kafka, Kafka Streams, and Apache Flink.
  • Experience with public cloud environments such as AWS, Google Cloud, and Azure, plus Terraform.
  • Strong knowledge of software engineering practices, including testing, CI/CD, Jenkins, GitHub Actions, agile development, Git/version control, and containers.
  • Excellent communication and collaboration skills, with the ability to work effectively with cross-functional teams.
  • Passion for staying current with emerging data engineering technologies and trends.

Skills

Apache Airflow, Dagster, dbt, Spark, SQL, Amazon Athena, Trino, Amazon Redshift, Snowflake, Postgres, MySQL, Apache Kafka, Apache Flink, AWS, GCP

Vanta

Vanta

Remote

Staff Software Engineer, Foundations
$260k+/yrRemote7+ YOEData Engineering

Leads the re-platforming of Vanta’s compliance data layer from MongoDB to schema-aware PostgreSQL across high-throughput Kafka and S3 pipelines. The role requires staff-level distributed systems expertise, migration leadership, and strong experience with relational and document data modeling.

Intercom

Intercom

Dublin, Ireland
Staff Data Engineer - GTM
No salary listedHybrid7+ YOEData Engineering

Builds the account, contact, hierarchy, enrichment, and identity systems that power go-to-market operations. The role requires modern data-stack experience, production LLM development, entity resolution, SaaS integrations, and close partnership with Sales, Marketing, and RevOps.

Altana

Altana

London, United Kingdom

Staff Implementations Data Engineer
No salary listedOn-site8+ YOEData Engineering

Build and scale distributed data ingestion, standardization, and big-data pipelines while improving platform reliability, security, and cost controls. The role requires 8+ years of data engineering experience, technical leadership, cloud deployment expertise, and strong customer-facing communication.

Alpaca

Alpaca

Remote

Senior Data Engineer
No salary listedRemote5+ YOEData Engineering

Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.

Elliptic

Elliptic

London, United Kingdom

Senior Software Engineer - Data
No salary listedHybrid5+ YOEData Engineering

Designs and builds scalable distributed data systems and pipelines powering blockchain analytics products. The role is hands-on, requiring production experience with Scala, Java, or Python, big-data technologies, cloud infrastructure, and data orchestration, plus technical leadership and mentoring.