Backend Engineer, Data
Build and own data pipelines, models, marts, and services supporting Product, Data Science, and go-to-market teams. The role requires 6+ years of software engineering experience, distributed data processing expertise, backend development skills, and strong SQL.
About the job
Responsibilities
- Design, develop, and own data pipelines, models, and products that power Product, Data Science, and go-to-market functions.
- Develop subject matter expertise and manage service-level agreements for data pipelines and full-stack web applications supporting critical stakeholders.
- Build and refine data foundations, including infrastructure, pipelines, and tools, using Scala, Spark, and Airflow.
- Leverage LLMs and agents at scale to produce high-quality data for ambiguous problems.
- Refine data marts that help the go-to-market organization forecast business performance and measure attainment toward targets.
- Build data services that track key product metrics and measure the impact of field-team strategies.
Requirements
- 6+ years of experience in software engineering, focused on building and maintaining data services or data-intensive applications.
- Strong engineering background and interest in data.
- Experience writing and debugging data pipelines using a distributed data framework such as Spark, Hadoop, or Pig.
- Ability to investigate data inconsistencies and resolve deep-rooted data quality issues.
- Knowledge of a backend development language such as Scala, Java, or Go.
- Strong SQL experience.
- Ability to communicate cross-functionally, derive requirements, and architect shared datasets.
Nice-to-haves
- Experience creating and maintaining data marts for business reporting.
- Experience working with Product or go-to-market teams, including Sales and Marketing.
Skills
Scala, Spark, Apache Airflow, Java, SQL, Python, Go, Hadoop, Pig, LLMs, Data Pipelines, Data Marts, Data Services, Data Quality, Full-Stack Web Applications
Similar jobs
Data Engineering jobsBuild and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.
Senior Data Engineer responsible for building and operating reliable clinical and claims data pipelines, CDC systems, quality controls, and de-identified exports. The role requires 5+ years of production pipeline experience plus strong SQL, Python, Spark, and data-governance skills.
Leads architecture for offline experimentation and route simulation while building reliable, scalable data pipelines and backend services. The role requires 5+ years of backend or data engineering experience, distributed-systems expertise, strong SQL and Spark skills, and proficiency with modern cloud infrastructure.
Builds agentic AI, automated data workflows, and BI solutions for complex telecommunications datasets. The role requires 5+ years of technical data and automation experience, strong SQL and Python skills, and expertise in data governance and LLM-based tools.
The Senior Data and AI Specialist will build agentic AI solutions, automated Python workflows, and analytics products across telecommunications data. The role requires at least five years of experience with SQL, data automation, BI tools, LLM agents, and data governance.