Senior Data Architect (USA)
Designs and leads unified data architecture integrating vendor datasets for quantitative research, simulation, and alpha generation across asset classes. Requires 7+ years experience in data engineering, Python proficiency, and financial data modeling expertise.
About the job
Responsibilities
- Architect and implement a unified data platform that integrates hundreds of vendor datasets, providing consistent, accessible, and high-quality data to simulators and researchers.
- Design efficient storage and retrieval systems to support both large-scale historical backtesting and high-frequency research workflows.
- Develop intuitive researcher interfaces and APIs that allow users to easily discover variables, explore metadata, and assemble data into standardized stocks × values matrices for rapid hypothesis testing.
- Collaborate closely with quantitative researchers and simulation teams to understand their workflows, ensuring the data platform meets real-world analytical and performance needs.
- Establish best practices for data modeling, normalization, versioning, and quality control across asset classes and data vendors.
- Work with infrastructure and DevOps teams to optimize data pipelines, caching, and distributed storage for scalability and reliability.
- Prototype and deploy internal data applications that enhance research productivity and data transparency.
- Mentor and guide data engineers to maintain robust, maintainable, and well-documented data systems.
Requirements
- 7+ years of experience in data architecture, quantitative research infrastructure, or large-scale data engineering in a financial or research-driven environment.
- Proven experience designing and implementing scalable data storage solutions (e.g., columnar databases, time-series systems, object stores, or data lakes).
- Strong proficiency in Python and familiarity with modern data stack technologies (e.g., Parquet, Arrow, Spark, SQL/NoSQL, distributed file systems).
- Deep understanding of time-series and financial data modeling, including handling multiple vendors, instruments, and frequencies.
- Experience building data interfaces, APIs, or tools that serve researchers, data scientists, or quantitative analysts.
- Ability to translate research needs into efficient data schemas and access patterns.
- Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, Mathematics, or a related quantitative field.
- Strong collaboration, communication, and documentation skills.
Nice-to-Haves
- Familiarity with cloud-based architectures (e.g., AWS, GCP, Azure) and modern data governance practices.
Compensation & Benefits
- Base salary range: $175,000 - $200,000 depending on candidate’s background.
- Bonus based on individual and company performance.
- PPO health, dental, and vision insurance premiums fully covered.
- Pre-tax commuter benefits.
- Weekly company meals.
Skills
Python, Parquet, Apache Arrow, Spark, SQL, NoSQL, AWS, GCP, Azure, Time-Series Databases
Similar jobs
Data Engineering jobsThe Senior Data Engineer will design scalable data pipelines and warehousing systems supporting analytics, business metrics, and machine-learning initiatives. The role requires 4+ years of enterprise data experience, expertise with modern data platforms and ETL, and the ability to mentor engineers and collaborate across functions.
The Senior Platform Engineer will build and operate reliable data platform tooling, consolidate orchestration, scale dbt infrastructure, and improve Databricks developer experience. The role requires 5+ years of production software experience, strong Python and AWS expertise, infrastructure-as-code experience, and familiarity with modern data stacks.
Senior software engineer building and evolving Fetch’s data platform, including pipelines, governed data access, delivery infrastructure, and partner integrations. The role requires 8+ years of experience, strong platform or backend expertise, ownership of complex cross-team initiatives, and excellent technical judgment.
Lead the development and maintenance of scalable data pipelines, warehouse, and transformation layer using modern data stack. Collaborate with data scientists and analysts to ensure clean, reliable data for insights in a high-growth startup.
Senior Data Engineer responsible for designing and operating scalable data pipelines and platform capabilities across Snowflake and AWS. The role requires 5+ years of production data engineering experience, strong SQL and Python skills, and expertise in ETL/ELT, orchestration, quality, and observability.