Skip to content
AnthropicAnthropic

Safeguards Enforcement Lead, Cyber Harms

Leads cyber-focused AI misuse enforcement, managing analysts and contractors while developing detection and mitigation strategies for attacks, malware, and exploitation. Requires people management, cybersecurity expertise, high-volume abuse enforcement, data analysis with SQL or Python, and cross-functional risk communication.

About the job

Responsibilities

  • Manage Cyber Enforcement Analysts and contractors, overseeing the vision and execution of cyber enforcement strategy.
  • Develop strategies to detect and mitigate misuse of AI systems for cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations.
  • Collaborate with stakeholders on novel, ambiguous, and high-severity cases.
  • Partner with the Safeguards Policy Design Team to address policy gaps identified through enforcement scenarios.
  • Work with Engineering and Data Science teams to support enforcement tooling and measurement.
  • Monitor emerging AI policy-enforcement practices, threat actor tactics, and the evolving cyber threat landscape.
  • Respond to escalations during weekends and holidays as needed.

Requirements

  • Experience managing people.
  • Cybersecurity experience, including offensive techniques, exploit development, malware analysis, or vulnerability research.
  • Experience with content review, abuse investigations, or high-volume policy enforcement.
  • Proficiency in SQL and/or Python for data analysis and threat detection.
  • Experience identifying emerging risks and communicating findings to Product, Policy, Engineering, and Legal stakeholders.
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement.
  • Bachelor's degree or equivalent combination of education, training, and experience in a relevant field.

Nice-to-haves

  • Experience in trust and safety, abuse investigations, cybersecurity investigations, or threat intelligence at a technology or AI company.
  • Experience with large language models and AI misuse in cyber operations.
  • Experience operating abuse-monitoring programs or enforcement-review systems.
  • Understanding of implementing product policies at scale, including content moderation.
  • Experience working with government agencies, regulated environments, or information-sharing communities.

Compensation

  • Annual salary: $285,000–$330,000 USD.

Skills

Cybersecurity, Exploit Development, Malware Analysis, Vulnerability Research, SQL, Python, Generative AI, Prompt Engineering, LLMs, Threat Intelligence, Content Moderation, Data Analysis

OpenAI

OpenAI

San Francisco, CA

Software Security Architect, Operating Systems | Consumer Devices
$268k+/yrOn-site7+ YOESecurity Engineering

Defines the security architecture for a next-generation operating system, spanning trust boundaries, hardware-backed protections, isolation, secure updates, and AI-agent guardrails. The role requires deep privileged-systems expertise, systems programming ability, and experience securing platforms across hardware, firmware, and software.

OpenAI

OpenAI

San Francisco, CA

Cyber Operations Lead, Critical Harm Operations
$252k+/yrHybrid8+ YOESecurity Engineering

Leads cybersecurity and cyber intelligence operations for high-risk user-safety decisions, combining strategic planning, operational systems, automation, and direct people management. Requires 8+ years in cybersecurity-related work and 4+ years leading teams.

Anthropic

Anthropic

San Francisco, CA
Platform Security Engineer, DRTM / Secure Launch
$320k+/yrHybrid8+ YOESecurity Engineering

Owns DRTM adoption, attestation, and platform hardening across x86 and ARM infrastructure, working across firmware, bootloaders, kernels, hardware, and silicon security. The role requires deep systems-security experience, upstream Linux or firmware contributions, and strong vendor and OEM leadership.

Vanta

Vanta

Remote

Lead Product GRC Subject Matter Expert
$230k+/yrRemote10+ YOESecurity Engineering

Leads interpretation and productization of federal compliance controls for Vanta’s public-sector platform, translating FedRAMP and related frameworks into technically testable guidance, automated detectors, mappings, and machine-readable authorization workflows. Requires 8–10+ years of hands-on federal compliance experience, especially FedRAMP program and SSP work.

Mercor

Mercor

San Francisco, CA

Security GRC Lead
$350k+/yrOn-site7+ YOESecurity Engineering

Leads the company’s security GRC function, owning SOC 2, ISO 27001, enterprise audits, third-party risk, policy governance, and automated evidence workflows. Requires 7+ years of GRC or audit experience, end-to-end SOC 2 and ISO 27001 ownership, and strong security tooling expertise.