Skip to content
AnthropicAnthropicSan Francisco, CA

Staff+ Software Engineer, Data Infrastructure

Build and scale secure data infrastructure at Anthropic, including access control systems, financial data pipelines, cloud storage reliability, and analytics tooling. Requires 10+ years building data/distributed systems, 3+ years leading complex projects, and deep experience with cloud/IaC or programming languages.

405k – 485k/yr
Hybrid10+ YOEData Engineering

About the role

Responsibilities

Within Data Infra, you may be matched to critical business areas including:

Data Governance & Access Control: Design and implement robust access control systems ensuring only authorized users can access sensitive data. Build infrastructure for permission management, audit logging, and compliance requirements. Work on IAM policies, ACLs, and security controls that scale across thousands of users and systems.

Financial Data Infrastructure: Build and maintain data pipelines and warehouses powering business-critical reporting. Ensure data integrity, accuracy, and availability for complex financial systems, including third party revenue ingestion pipelines; manage the external relationships as needed to drive upstream dependencies. Own the reliability of systems processing revenue, usage, and business metrics.

Cloud Storage & Reliability: Architect disaster recovery, backup, and replication systems for petabyte-scale data. Ensure high availability and durability of data stored in cloud object storage (GCS, S3). Build systems that protect against data loss and enable rapid recovery.

Data Platform & Tooling: Scale data processing infrastructure using technologies like BigQuery, BigTable, Airflow, dbt, and Spark. Optimize query performance, manage costs, and enable self-service analytics across the organization.

Requirements

  • 10+ years (not including internships or co-ops) of experience in a Software Engineer role, building data infrastructure, storage systems, or related distributed systems
  • 3+ years (not including internships or co-ops) of experience leading large scale, complex projects or teams as an engineer or tech lead
  • Can set technical direction for a team, not just execute within it
  • Deep experience with at least one of:
    • Strong proficiency in programming languages like Python, Go, Java, or similar
    • Experience with infrastructure-as-code (Terraform, Pulumi) and cloud platforms (GCP, AWS)
  • Can navigate complex technical tradeoffs between performance, cost, security, and maintainability
  • Excellent collaboration skills - you work well with both technical and non-technical stakeholders
  • Bachelor’s degree or an equivalent combination of education, training, and/or experience in a field relevant to the role

Nice-to-Haves

  • Experience with security and compliance requirements (ITGC, GDPR, financial controls)
  • Background in data warehousing, ETL/ELT pipelines, or analytics infrastructure
  • Experience with Kubernetes, containerization, and cloud-native architectures
  • Track record of improving data reliability, availability, or cost efficiency at scale
  • Knowledge of column-oriented databases, OLAP systems, or big data processing frameworks
  • Experience working in fintech, financial services, or highly regulated environments
  • Security engineering background with focus on data protection and access controls

Technologies

  • Data: BigQuery, BigTable, Airflow, Cloud Composer, dbt, Spark, Segment, Fivetran
  • Storage: GCS, S3
  • Infrastructure: Terraform, Kubernetes, GCP, AWS
  • Languages: Python, Go, SQL

Skills

PythonGoJavaTerraformPulumiGCPAWSKubernetesBigQuerySparkAirflowdbtSQL
Anthropic

Staff+ Software Engineer, Capacity Engineering

AnthropicSan Francisco, CA +2

Build and operate production data pipelines, observability tools, and planning systems to maximize utilization, efficiency, and attribution of Anthropic's large-scale multi-cloud accelerator and CPU fleet. Requires strong Python/SQL, cloud operations, and Kubernetes experience in a high-ambiguity environment.

320k – 485k/yr
Hybrid7+ YOEData Engineering
Anthropic

Staff+ Software Engineer, Databases

AnthropicSan Francisco, CA +2

Build and scale the core database infrastructure powering Claude at Anthropic, including data plane/control plane, data movement (CDC, migrations), and caching systems that support millions of users and frontier AI research across multi-cloud environments. Requires deep expertise in distributed databases and production storage systems.

320k – 485k/yr
Hybrid10+ YOEData Engineering
The Voleon Group

Staff Software Engineer, Batch and Realtime Streaming

The Voleon GroupBerkeley, CA +1

Architect and build Voleon’s batch and realtime streaming platform that powers ML research and production trading systems. Requires 10+ years experience building scalable data infrastructure with Python, Go, and distributed systems.

315k – 405k/yr
Remote10+ YOEData Engineering
Runway

Member of Technical Staff, Research Engineer (Datasets)

RunwayUnited States

Research Engineer owning datasets for training world simulation AI models, designing multimodal datasets, running experiments, and building data pipelines to enhance model capabilities across tasks like robotics and creative tools. Requires 4+ years in ML with experience in generative models and frameworks like PyTorch or JAX.

270k – 370k/yr
Remote4+ YOEData Engineering
Scale AI

Staff Software Engineer, Data Platform

Scale AISan Francisco, CA +2

Leads architecture and development of large-scale data platforms for AI, including storage, streaming, caching, and indexing. Requires 8+ years experience with databases, streaming tools, Kubernetes, and distributed systems.

248k – 311k/yr
Hybrid8+ YOEData Engineering