# Safeguards Enforcement Lead, Cyber Harms

**Company:** [Anthropic](https://hotfix.jobs/companies/anthropic)
**Location:** Washington, DC, San Francisco, CA, New York, NY
**Role:** Security Engineering
**Salary:** $285k – $330k/yr
**Skills:** Cybersecurity, Exploit Development, Malware Analysis, Vulnerability Research, SQL, Python, Generative AI, Prompt Engineering, LLMs, Threat Intelligence, Content Moderation, Data Analysis
**Posted:** 2026-08-29

> Leads cyber-focused AI misuse enforcement, managing analysts and contractors while developing detection and mitigation strategies for attacks, malware, and exploitation. Requires people management, cybersecurity expertise, high-volume abuse enforcement, data analysis with SQL or Python, and cross-functional risk communication.

## Job Description

## Responsibilities
- Manage Cyber Enforcement Analysts and contractors, overseeing the vision and execution of cyber enforcement strategy.
- Develop strategies to detect and mitigate misuse of AI systems for cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations.
- Collaborate with stakeholders on novel, ambiguous, and high-severity cases.
- Partner with the Safeguards Policy Design Team to address policy gaps identified through enforcement scenarios.
- Work with Engineering and Data Science teams to support enforcement tooling and measurement.
- Monitor emerging AI policy-enforcement practices, threat actor tactics, and the evolving cyber threat landscape.
- Respond to escalations during weekends and holidays as needed.

## Requirements
- Experience managing people.
- Cybersecurity experience, including offensive techniques, exploit development, malware analysis, or vulnerability research.
- Experience with content review, abuse investigations, or high-volume policy enforcement.
- Proficiency in SQL and/or Python for data analysis and threat detection.
- Experience identifying emerging risks and communicating findings to Product, Policy, Engineering, and Legal stakeholders.
- Experience working with generative AI products, including writing effective prompts for content review and enforcement.
- Bachelor's degree or equivalent combination of education, training, and experience in a relevant field.

## Nice-to-haves
- Experience in trust and safety, abuse investigations, cybersecurity investigations, or threat intelligence at a technology or AI company.
- Experience with large language models and AI misuse in cyber operations.
- Experience operating abuse-monitoring programs or enforcement-review systems.
- Understanding of implementing product policies at scale, including content moderation.
- Experience working with government agencies, regulated environments, or information-sharing communities.

## Compensation
- Annual salary: $285,000–$330,000 USD.

## Similar jobs

- [Software Security Architect, Operating Systems | Consumer Devices](https://hotfix.jobs/jobs/eae7f6bf-7cf9-495a-9c46-fb88a4620a0f) - OpenAI - San Francisco, CA - $268k – $342k/yr
- [Cyber Operations Lead, Critical Harm Operations](https://hotfix.jobs/jobs/aece8b0b-4900-4762-87e1-025b2cbe75c9) - OpenAI - San Francisco, CA - $252k – $335k/yr
- [Platform Security Engineer, DRTM / Secure Launch](https://hotfix.jobs/jobs/52321fb7-76c2-4b9b-b900-53a63bb68154) - Anthropic - San Francisco, CA - $320k – $405k/yr
- [Lead Product GRC Subject Matter Expert](https://hotfix.jobs/jobs/57c937d5-05e3-4033-875a-890645c4aa6b) - Vanta - Remote - $230k – $270k/yr
- [Security GRC Lead](https://hotfix.jobs/jobs/2eb261b0-5ff9-435d-b0fb-eaf5933051d1) - Mercor - San Francisco, CA - $350k – $425k/yr

**Apply:** https://hotfix.jobs/jobs/c042fb99-fa5c-4740-9908-e14324222bde
**Canonical:** https://hotfix.jobs/jobs/c042fb99-fa5c-4740-9908-e14324222bde