Senior Software Engineer building large-scale distributed data systems and pipelines on Scala, Spark, and Databricks to power blockchain analytics and intelligence products that fight financial crime. Requires hands-on Scala/functional programming experience and expertise designing scalable batch/streaming data platforms in the cloud.
165k – 305k/yr
Hybrid5+ YOEData Engineering
About the role
Responsibilities
Write, ship, and maintain production code as a hands-on engineer.
Architect, design, and implement large-scale distributed data systems and pipelines.
Contribute to technical decision-making across batch and streaming data solutions.
Collaborate with engineers, product managers, data scientists, and intelligence analysts to build customer-focused products.
Explore and integrate new technologies (e.g. data orchestration or cloud-native tools) to optimise performance and scalability.
Take shared ownership of data systems, from design to deployment and ongoing improvement and support.
Perform thoughtful peer reviews that raise code quality and share best practices across the team.
Contribute to platform-wide initiatives that improve reliability, observability, and cost efficiency.
Help shape the technical roadmap for data engineering across Elliptic.
Leverage and deploy AI and agentic systems to operate more effectively.
Mentor and upskill engineers, champion best practices, and hold the team to a high standard.
Requirements
Hands-on production experience with Scala and functional programming.
Ability to design, build, and maintain distributed data systems in a cloud-based environment.
Hands-on experience with big data frameworks such as Spark or Databricks.
Knowledge of cloud infrastructure (AWS, GCP, or Azure).
Judgement to balance scalability, performance, and maintainability.
Experience with data modelling and workflow orchestration.
Nice-to-Haves
Experience in stream processing frameworks and event-driven architecture.
Hands-on experience with Infrastructure as Code (Terraform, CloudFormation).
Experience working in containerised environments (Docker, Kubernetes).
Interest in AI-driven tooling for engineering workflows.
Designs and runs massive-scale data pipelines for ingestion, normalization, enrichment, and delivery across 80M+ companies and 800M+ people. Manages data operations, BPO vendors, partnerships, monitoring, and cost optimization using Python, Dagster, and DuckDB.
165k – 250k/yrOn-siteData Engineering
Manager, Analytics Engineering
JustworksNew York, NY
Lead and develop a team of analytics engineers to design, build, and maintain scalable data models, ELT pipelines, and BI solutions using modern data stack tools. Requires 7+ years data experience including 2+ years managing teams, deep expertise in SQL, Python, Snowflake, dbt, and dimensional modeling.
166k – 214k/yrHybrid7+ YOEData Engineering
Senior Software Engineer, Data Engineering
ChimeSan Francisco, CA
Builds and scales ETL pipelines, designs data schemas, and owns data quality/governance for 10x growth. Requires 5+ years in data pipelines with SQL, Spark, Airflow, Python, and MPP databases like Snowflake/Redshift.
164k – 227k/yrHybrid5+ YOEData Engineering
Senior Software Engineer - Distributed Data Systems
DatabricksBellevue, WA +2
Senior engineer building distributed data systems like Apache Spark and Delta Lake to handle big data processing, ETL, and data science workloads. Requires 5+ years in Java/Scala/C++ and expertise in distributed systems.
166k – 225k/yrOn-site5+ YOEData Engineering
Senior Software Engineer - Distributed Data Systems
DatabricksMountain View, CA
Develop distributed data systems including Apache Spark and Delta Lake to handle big data workloads efficiently. Requires 5+ years in Java/Scala/C++ and expertise in distributed systems.