[P] Data Engineer, Safeguards
Build and maintain scalable data pipelines, warehousing, dashboards, and analytical tooling to support Anthropic's Safeguards team in monitoring AI models, detecting abuse, and ensuring safety at scale. Requires strong SQL/Python, modern data stack experience, and collaboration with engineers, data scientists, and policy teams.
About the job
Key Responsibilities
- Design, build, and maintain scalable data pipelines that support safety monitoring, abuse detection, and enforcement workflows
- Develop and optimize data models and warehousing solutions to enable efficient analysis of large-scale usage and safety data
- Build and maintain dashboards and reporting infrastructure that give Safeguards teams visibility into model behavior, misuse patterns, and enforcement outcomes
- Collaborate with engineers to integrate data from multiple sources — including model outputs, user reports, and automated classifiers — into a unified analytical layer
- Implement data quality frameworks, monitoring, and alerting to ensure the reliability of safety-critical data
- Partner with research teams to surface data insights that inform model improvements and safety interventions
- Develop self-service data tooling that enables stakeholders to explore safety data and generate reports independently
- Contribute to data governance practices, including access controls, retention policies, and privacy-compliant data handling
Minimum Qualifications
- Proficiency in SQL and Python, with hands-on experience building and maintaining ETL/ELT pipelines
- Experience with cloud data platforms such as BigQuery, Redshift, Snowflake, or similar
- Experience with modern data stack tools such as dbt, Airflow, Spark, or similar orchestration and transformation frameworks
- Experience building dashboards and data visualizations using tools such as Looker, Tableau, or Metabase
- Ability to communicate clearly and translate complex data concepts for both technical and non-technical audiences
Preferred Qualifications
- 8+ years of experience in data engineering, analytics engineering, or a related role
- Comfort contributing across the stack and picking up work outside your immediate scope when the situation calls for it
- Background in trust and safety, integrity, fraud, or abuse detection data systems
- Experience with large-scale event streaming systems such as Kafka, Pub/Sub, or Kinesis
- Experience building data infrastructure that supports ML model monitoring or evaluation
- Familiarity with data privacy and compliance frameworks such as GDPR, CCPA, or similar
- Background in statistical analysis or experience working closely with data scientists
- A genuine interest in the societal implications of AI and in making AI systems safer
Compensation
Annual Salary: $320,000—$405,000 USD
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Skills
SQL, Python, ETL, ELT, BigQuery, Redshift, Snowflake, dbt, Airflow, Spark, Looker, Tableau, Metabase, Kafka
Similar jobs
Data Engineering jobsLeads the architecture, scaling, security, governance, and cost optimization of enterprise and AI data platforms. Requires 10+ years of data or software engineering experience, with expertise in production data foundations, CI/CD, governance, security, and performance optimization.
Build and scale data pipelines, reusable datasets, and validation frameworks supporting business intelligence, marketing, and data science. The role requires strong Python and SQL skills, modern data-stack experience, and at least four years of software or data engineering experience.
Own the reliability, performance, observability, scalability, and cost efficiency of large Aurora MySQL production environments supporting healthcare applications. The role requires 6+ years of database engineering experience, deep MySQL and AWS expertise, and strong skills in automation, incident response, and query optimization.
Leads an analytics engineering team that transforms raw data into reliable, actionable insights for product, marketing, and operations. The role requires 7+ years in data or analytics engineering, management experience, and advanced SQL, Databricks, and dbt expertise.
Senior Data Infrastructure Engineer responsible for building and operating reliable, low-latency streaming and batch data systems that support AI products. Requires 5+ years of production data infrastructure experience and expertise with technologies such as Kafka, Flink, ClickHouse, and Terraform.