# Red Team Specialist - Cyber

**Company:** [OpenAI](https://hotfix.jobs/companies/openai)
**Location:** San Francisco, CA, Seattle, WA, Washington, DC
**Role:** Security Engineering
**Salary:** $198k – $320k/yr
**Skills:** Cybersecurity, Application Security, Penetration Testing, Vulnerability Research, Adversary Simulation, Red Teaming, Python, Ai Model Evaluation, Agentic Systems, Prompt Injection, Jailbreaking, Automation, Statistical Analysis, Security Testing, Risk Assessment
**Posted:** 2026-08-18

> The Red Team Specialist evaluates AI models for cyber capabilities, safeguard failures, and agentic-system abuse risks. The role combines hands-on security testing, automated evaluation infrastructure, risk assessment, and cross-functional communication.

## Job Description

## Responsibilities
- Design and run rigorous evaluations of model cyber capabilities and safeguards, including policy adherence, correct refusal, over-refusal, and resilience to jailbreaking and other adversarial techniques.
- Conduct hands-on testing to understand what models can enable when used by experienced security practitioners, including through task-specific harnesses, scaffolding, and multi-step workflows.
- Distinguish benchmark or policy failures from behavior that creates meaningful real-world risk, considering feasibility, attacker uplift, reliability, and capabilities already available elsewhere.
- Build and improve automated testing infrastructure for repeatable measurement, rapid iteration, and statistically grounded analysis across models and product surfaces.
- Test novel abuse risks in agentic systems, including indirect prompt injection, agent hijacking, and other adversarial manipulation of tool-using systems.
- Translate findings into clear risk assessments and actionable recommendations for Security, Research, Product, Policy, and Engineering partners.
- Contribute to Safety Bug Bounty work, particularly reports requiring cyber expertise.

## Requirements
- Substantial depth in at least one of the following areas:
  - Cybersecurity, including application security, penetration testing, vulnerability research, adversary simulation, or red-team operations.
  - AI model evaluation, including designing and running evaluations, building agentic harnesses, automating adversarial testing, constructing datasets, or analyzing model behavior at scale.
- Working literacy across both cybersecurity and model evaluation, with interest in developing further depth outside the primary area.
- Ability to write code and build practical testing tools, particularly for automating experiments, orchestrating models, or analyzing results.
- An attacker mindset and interest in discovering failure modes that standard evaluations may not capture.
- Clear written and verbal communication, including the ability to explain technical findings, limitations, and risk to audiences with different backgrounds.
- Experience working across technical and non-technical teams to move from a finding to a decision, mitigation, or follow-up test.

## Compensation
- Salary range: $198,000–$320,000.

## Similar jobs

- [Software Engineer, Trust & Safety](https://hotfix.jobs/jobs/ccbdc9a4-a902-479b-845f-0d13a50d97a0) - Vercel - San Francisco, CA - $196k – $294k/yr
- [Manager, Security Incident Response](https://hotfix.jobs/jobs/fe4d4111-7237-4acc-be44-fce689acc71f) - 1Password - Remote - $192k – $278k/yr
- [Governance, Risk, and Compliance Manager - Privacy](https://hotfix.jobs/jobs/a5153f89-2b1c-40d2-b939-13b82952a049) - Decagon - San Francisco, CA - $190k – $275k/yr
- [Safeguards Policy Analyst, Cyber Harms](https://hotfix.jobs/jobs/b7a7ba79-b5ba-42b9-930f-19f379ea9abd) - Anthropic - San Francisco, CA - $190k – $285k/yr
- [Governance, Risk, and Compliance Manager](https://hotfix.jobs/jobs/5e407075-824a-4639-96f7-821772a6f497) - Decagon - San Francisco, CA - $190k – $275k/yr

**Apply:** https://hotfix.jobs/jobs/9bb5b4f5-abe4-4fc7-8226-8810777d7d5d
**Canonical:** https://hotfix.jobs/jobs/9bb5b4f5-abe4-4fc7-8226-8810777d7d5d