Senior Software Engineer, Strategy Research Analytics
Leads design and evolution of analytics infrastructure for research reporting, owning pipelines, datasets, and platform standardization. Collaborates with data scientists using Python, SQL, distributed engines like Presto/Spark. Requires 6+ years in data infrastructure.
About the job
Responsibilities
- Own implementation and on-going operation of recurring analytics pipelines (e.g., Airflow DAGs) including monitoring, alerting, and reliability improvements
- Lead architectural evolution of the analytics platform, including schema standardization, DAG consolidation, and modernization of legacy workflows
- Drive cross-team technical alignment when consolidating duplicated or inconsistent analytics outputs
- Build and maintain base analytics tables and metrics with strong schema discipline and reproducible computation
- Define and implement reliability standards (SLOs, observability patterns, runbooks) adopted across analytics pipelines
- Improve transparency and usability through documentation, discoverability, and clear data contracts
- Optimize distributed compute and SQL query performance; design data layouts (partitioning, file sizing) for columnar storage
Requirements
- Bachelor's degree in Computer Science or equivalent professional experience
- 6+ years of experience building and operating analytics or data infrastructure systems
- Strong proficiency in Python and SQL
- Deep experience with distributed query engines and large-scale compute systems
- Demonstrated ownership of large-scale or mission-critical data infrastructure
- Strong data modeling expertise, including schema design, partitioning strategy, and reproducibility considerations
- Expertise in metadata management, data lineage, and applying robust data governance principles
Preferred Qualifications
- Experience leading architectural migrations or major refactors of data platforms
- Familiarity with AWS cloud technologies and on-prem compute clusters (e.g., Slurm, SSH, Unix)
- Exposure to quantitative research or machine learning environments
Skills
Python, SQL, Airflow, Presto, Spark, Parquet, Orc, AWS, Data Modeling, Data Lineage, Metadata Management, Slurm
Similar jobs
Data Engineering jobsSenior Data Infrastructure Engineer responsible for building and operating reliable, low-latency streaming and batch data systems that support AI products. Requires 5+ years of production data infrastructure experience and expertise with technologies such as Kafka, Flink, ClickHouse, and Terraform.
Build and operate petabyte-scale data infrastructure powering Discord’s insights and products. The role requires 5+ years of software engineering experience, strong programming skills, and experience with large-scale pipelines, streaming, orchestration, or data warehousing.
Build and lead the central data platform, covering ingestion, warehousing, orchestration, streaming, self-service frameworks, and trust layers. The role requires 5+ years of production data infrastructure experience, strong Python and SQL skills, and expertise with Snowflake and modern data tooling.
Build and operate large-scale revenue data pipelines powering billing and cost attribution, while improving reliability, latency, and correctness. The role requires strong Spark and Airflow experience, cross-functional problem-solving, and operational ownership of mission-critical production systems.
Build and operate low-latency systems that capture, normalize, and distribute real-time market data for institutional trading. The role requires backend engineering experience, Java or C++, market data infrastructure knowledge, and exchange connectivity expertise.