Skip to content
OpenAIOpenAI

Data Engineer, Monetization Data Platform

Build and operate scalable monetization data platforms, pipelines, models, and quality systems spanning product, financial, and operational data. The role partners with Product Engineering, Finance, Accounting, Analytics, and GTM teams to deliver reliable, observable data products.

About the job

Responsibilities

  • Design, build, and operate large-scale streaming and batch data pipelines for product, financial, and operational data.
  • Develop canonical data models and reusable data products for product usage, pricing, billing, ads, payments, revenue, and general-ledger domains.
  • Establish guarantees for data accuracy, completeness, freshness, lineage, reconciliation, and auditability.
  • Build platform capabilities and frameworks that improve developer productivity and enable trusted monetization data products.
  • Partner with Product Engineering, Finance, Accounting, Analytics, and GTM teams to define data contracts and instrument monetization features.
  • Lead technical design and delivery of complex cross-functional projects using system designs and RFCs.
  • Improve observability and operational excellence through monitoring, incident response, root-cause analysis, and remediation.
  • Contribute to engineering standards, documentation, and knowledge sharing.

Requirements

  • Deep experience building and operating production data platforms, distributed data systems, or high-scale data pipelines.
  • Proficiency in large-scale data pipeline architecture and at least one general-purpose programming language such as Python, Java, or Scala.
  • Strong fundamentals in data modeling, data architecture, distributed systems, and software engineering.
  • Experience designing systems with rigorous data quality, observability, lineage, governance, privacy, or access-control requirements.
  • Ability to collaborate with technical and non-technical partners, navigate ambiguity, and deliver scalable technical solutions.
  • Strong communication, engineering judgment, and focus on correctness and operational reliability.

Nice to Have

  • Experience with monetization, pricing, product usage, billing, ads, payments, revenue, or financial data.
  • Familiarity with financial controls, reconciliation, close processes, or audit requirements.
  • Experience with lakehouse or data warehouse technologies, workflow orchestration, streaming systems, and data transformation frameworks.
  • Experience building self-service data platforms, shared frameworks, or developer tooling for data and engineering teams.

Compensation

  • Salary range: $230,000–$385,000.

Skills

Python, Java, Scala, Data Pipelines, Data Modeling, Data Architecture, Distributed Systems, Data Quality, Observability, Data Lineage, Data Governance, Workflow Orchestration, Streaming Systems, Data Warehousing

The Voleon Group

The Voleon Group

New York, NY
Software Engineer, Strategy Research Analytics
$230k+/yrRemote3+ YOEData Engineering

Build and evolve reliable analytics infrastructure, pipelines, schemas, and foundational datasets supporting quantitative research across strategies. The role requires strong Python and SQL skills, distributed data-platform experience, and ownership of observability, performance, and reproducibility.

Anyscale

Anyscale

San Francisco, CA

Software Engineer
$215k+/yrOn-site3+ YOEData Engineering

Build and optimize Ray Data, a Python-native data processing engine for large-scale AI workloads. The role focuses on distributed systems performance, scalable data pipelines, production training solutions, and fault tolerance while partnering with AI-focused customers.

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Data Operations
$250k+/yrHybridData Engineering

Own end-to-end data sourcing and vendor operations that help researchers train and evaluate frontier AI models. The role requires strong judgment, communication, problem-solving, and comfort managing ambiguous, fast-changing projects.

Applied Intuition

Applied Intuition

Sunnyvale, CA

Data Engineer - Axion
$200k+/yrOn-site5+ YOEData Engineering

Builds scalable data pipelines and data engine architecture for machine learning, integrating foundation models to automate labeling and discovery. The role requires 5+ years of experience, modern ML infrastructure expertise, and U.S. citizenship with security-clearance eligibility.

Fluidstack

Fluidstack

Austin, TX
Data Engineer
$269k+/yrOn-site5+ YOEData Engineering

Build and own production data pipelines, knowledge graph data models, and structured datasets from messy sources (PDFs, spreadsheets, telemetry) to power internal tools, dashboards, and ML models at a frontier AI compute infrastructure company. Requires experience operating depended-on pipelines, schema modeling, data quality engineering, and unstructured data extraction.