Skip to content
11x11x

Data Engineer

Own and extend customer data ingestion platform and large-scale pipelines powering AI workers. Build data lake, retrieval layer, and infrastructure for syncing, enriching, and querying customer data across CRMs and third-party systems.

About the job

What You'll Do

  • Own and extend our customer data ingestion platform
  • Build and maintain large-scale data pipelines powering AI products and customer workflows
  • Design systems for syncing customer data across external platforms, CRMs, and third-party systems
  • Help architect our future data lake, retrieval layer, and data infrastructure strategy
  • Build ingestion and querying systems for lead, account, enrichment, and customer knowledge data
  • Create the infrastructure that gives our AI workers access to the information they need to reason, act, and improve over time
  • Partner closely with product and engineering teams to unlock new AI product capabilities
  • Improve reliability, observability, performance, and scalability across our data stack
  • Contribute across backend systems and infrastructure, not just traditional data engineering projects
  • Push ideas into production quickly instead of over-optimizing in planning phases

What We're Looking For

  • 4+ years of software engineering or data engineering experience
  • Strong experience building and maintaining production-grade data systems and pipelines
  • Experience with Python and Typescript
  • Strong backend engineering fundamentals beyond traditional data engineering
  • High agency — you naturally move things forward without waiting for direction
  • Comfort operating in ambiguity and building without a playbook
  • Strong systems thinking and architectural intuition
  • Ability to balance speed with long-term scalability
  • Strong problem-solving skills and a willingness to own problems end-to-end
  • Hungry to grow. You want expanding scope, responsibility, and ownership over time
  • Excitement about AI-native workflows and the future of software development

Nice to Have

  • Experience with ClickHouse or other columnar databases
  • Experience building customer data platforms (CDPs)
  • Experience with Airbyte or similar data integration and ingestion platforms
  • Experience with CRM integrations and synchronization systems
  • Experience designing data lake architecture
  • Experience supporting AI or ML products
  • You've used tools like Claude Code, Codex, Cursor, or similar heavily in your workflow
  • You care deeply about engineering velocity and iteration speed

Skills

Python, TypeScript, Data Pipelines, Data Ingestion, Data Lake Architecture, ClickHouse, Airbyte, Crm Integrations, Backend Engineering, Observability

Imprint

Imprint

New York, NY
Infrastructure Engineer
$170k+/yrOn-site5+ YOEData Engineering

Build and operate scalable data infrastructure, including partner data sharing, identity graph foundations, and governed batch and real-time platforms. The role requires 5+ years of data, distributed systems, infrastructure, or backend engineering experience and strong cloud and data-platform expertise.

Vanta

Vanta

Remote

Operations Manager, Signal Systems
$176k+/yrRemoteData Engineering

Own the systems that ingest, standardize, validate, and operationalize data signals for Vanta’s EPD organization. The role suits a hands-on builder who has recently shipped working tools or pipelines, uses AI-assisted development, and helps teammates grow technically.

xAI

xAI

Palo Alto, CA

Analytics Engineer - X
$180k+/yrOn-site4+ YOEData Engineering

Build scalable data pipelines, infrastructure, and quantitative models that support experimentation, forecasting, and business decision-making. The role requires 4+ years of production data engineering experience, strong Python and SQL skills, distributed computing expertise, and a quantitative degree.

Abridge

Abridge

San Francisco, CA

Data Engineer
$185k+/yrHybrid5+ YOEData Engineering

Builds and optimizes scalable data pipelines, storage, and OLAP databases for ML training, analytics, and product features. Requires 5+ years in data engineering, proficiency in Python/SQL/cloud platforms, and distributed systems experience.

Benchling

Benchling

San Francisco, CA

Data Engineer
$153k+/yrHybrid3+ YOEData Engineering

Build and operate reliable, production-grade data pipelines, warehouse infrastructure, and trusted datasets supporting company-wide analytics and AI initiatives. The role requires 3+ years of production data engineering experience, strong SQL and Python skills, and experience with Snowflake, dbt, cloud infrastructure, and orchestration.