Safeguards Enforcement Analyst, Cyber Harm
Review flagged content and execute enforcement actions to detect and mitigate misuse of Anthropic's AI for cyberattacks, malware, and cyber exploitation. Requires cybersecurity experience, content moderation at volume, SQL/Python proficiency, and working with generative AI products.
About the job
Key Responsibilities
- Review flagged content and accounts to make accurate, well-documented enforcement decisions in line with our usage policies
- Detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations
- Triage and escalate novel, ambiguous, or high-severity cases to appropriate stakeholders
- Provide detailed feedback to the Safeguards policy design team on policy gaps surfaced through real enforcement scenarios
- Partner with Engineering and Data Science teams by surfacing detection model errors and quality signals from review to improve precision and recall
- Maintain high accuracy and consistency standards across review queues
- Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions
Minimum Qualifications
- Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research
- Experience performing content review, abuse investigations, or policy enforcement at volume
- Proficiency in SQL and/or Python for data analysis and threat detection
- Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams
- Experience working with generative AI products, including writing effective prompts for content review and enforcement
Preferred Qualifications
- Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company
- Experience with large language models and an understanding of how AI technology could be misused for cyber operations
- Experience operating within abuse monitoring programs or enforcement review systems
- Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
- Experience working with government agencies, regulated environments, or information sharing communities
Education
- Bachelor’s degree or an equivalent combination of education, training, and/or experience in a field relevant to the role
Note: This role may involve exposure to explicit, violent, technical, or psychologically disturbing content and may require responding to escalations during weekends and holidays.
Skills
Cybersecurity, Exploit Development, Malware Analysis, Vulnerability Research, Content Review, Abuse Investigations, Policy Enforcement, SQL, Python, Threat Detection, Generative AI, LLMs, Trust & Safety, Threat Intelligence
Similar jobs
Build and run the global spares sourcing program for critical datacenter equipment (chillers, generators, switchgear) at multi-GW scale. Qualify/dual-source suppliers, set stocking levels tied to failure data and criticality, negotiate VMI/consignment terms, and partner with reliability teams to ensure zero downtime from parts shortages.
Own state and local policy agenda for AI data center infrastructure. Build relationships with officials and regulators while drafting testimony and navigating permitting and utility approvals.
Own public affairs and community relations for Fluidstack's data center and power infrastructure projects. Build local support through town halls and coalitions, manage media and opposition responses, and partner with internal teams to align external messaging with project realities.
Leads child safety investigations, enforcement decisions, mandatory-reporting workflows, and process improvements involving sensitive content and abuse signals. The role requires Trust & Safety investigation experience, sound judgment, strong documentation, and cross-functional collaboration.
Own and maintain Primavera P6 CPM schedules, cost tracking, and productivity reporting for greenfield hyperscale data center construction projects. Drive critical path compression, lead reviews with GCs/executives, and deliver defensible as-built records on $500M+ programs.