# Safety Operations Lead

**Company:** [Thinking Machines Lab](https://hotfix.jobs/companies/thinking-machines-lab)
**Location:** San Francisco, CA
**Role:** Security Engineering
**Salary:** $190k – $300k/yr
**Experience:** 7+ years
**Skills:** trust and safety, content moderation, fraud prevention, abuse detection, safety policy, prompt injection, jailbreak detection, Cybersecurity, AI Tools, LLM APIs
**Posted:** 2026-08-04

> Leads operational trust and safety for human-AI collaboration products by reviewing abuse cases, enforcing policy, and building automation and detection systems. Requires 7+ years of recurring case-queue experience plus expertise in AI abuse risks, policy operations, and production safety incidents.

## Job Description

## Responsibilities
- Review flagged content, safety escalations, and account-level abuse signals daily.
- Triage cases, apply policy judgment, and take action through content flags, account reviews, bans, and recovery workflows.
- Design and refine safety policy across the product stack with engineering, legal, safety research, and security stakeholders.
- Build and maintain tooling and automation, including triage agents, ban/recovery workflows, abuse detection frameworks, and templates.
- Partner with product teams to embed safety into product experiences, including model refusals, content flagging, account review, and safety protections.
- Improve observability and detection for safety-relevant events, model safety trends, abuse patterns, and malicious behavior in production.

## Requirements
- 7+ years in an operational trust and safety, content moderation, or fraud/abuse operations role with direct, recurring responsibility for a case queue.
- Experience owning policy definition, operationalization, and enforcement end to end.
- Direct case experience with cybersecurity abuse, CBRN-relevant risk, youth safety, or prompt injection in a production environment.
- Working familiarity with model safety and abuse risk categories, including jailbreaks, prompt injection, and scaled abuse.
- Practical experience using AI tools such as Claude, Codex, or similar to build or accelerate operational workflows.

## Nice-to-haves
- Experience with safety and integrity operations on AI-powered products or LLM APIs.
- Track record of turning recurring case patterns into reusable tooling, workflows, or process improvements while continuing to own the underlying queue.
- Experience training, onboarding, or setting quality standards for moderators or reviewers.

## Compensation and Benefits
- Expected annual salary range: **$190,000-$300,000**.
- Health, dental, and vision benefits.
- Unlimited PTO.
- Paid parental leave.
- Relocation support as needed.
- Visa sponsorship.

## Similar roles

- [Senior Information Security Engineer](https://hotfix.jobs/jobs/ce08603c-ef42-46b3-9e93-601ca9a1e56f) - Zoox - Foster City, CA - $190k – $228k/yr
- [Senior Application Security Engineer](https://hotfix.jobs/jobs/45402dbe-ee8f-4a25-9817-c293c712505a) - Apollo - Remote - $190k – $273k/yr
- [Senior Software Engineer - Security](https://hotfix.jobs/jobs/e3185b76-5a92-4a24-956c-5a19fd20d75d) - Skydio - San Mateo, CA - $190k – $250k/yr
- [Senior Security Engineer, Application & Platform Security](https://hotfix.jobs/jobs/935a9604-4c7b-48d9-a08b-cb94b61fb298) - Sentry - San Francisco, CA - $190k – $280k/yr
- [Senior Software Engineer, Anti-Abuse & Security](https://hotfix.jobs/jobs/af883f53-3346-4cc3-827e-3e9664ff95c3) - Replit - Foster City, CA - $190k – $240k/yr

**Apply:** https://hotfix.jobs/jobs/c83ceda2-d7a3-405c-ae10-b680f0dd7577
**Canonical:** https://hotfix.jobs/jobs/c83ceda2-d7a3-405c-ae10-b680f0dd7577