# Safeguards Enforcement Analyst, Ban Evasion & Recidivism

**Company**: Anthropic
**Location**: Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC
**Work arrangement**: remote
**Experience**: senior
**Job type**: full-time
**Salary**: $245,000-$285,000 USD
**Category**: Operations
**Industry**: Technology
**Wikidata**: https://www.wikidata.org/wiki/Q116758847

**Apply**: https://job-boards.greenhouse.io/anthropic/jobs/5319592008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_dd4ec94e-48d

## Description

As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm.

Your initial focus will be recidivism: a ban that an actor can evade in five minutes isn't enforcement , it's friction. You'll own detecting when banned actors return, linking accounts across identities, and closing the re-registration paths that matter most. The mandate includes our highest-stakes populations, including preventing evasion of child-safety enforcement bans, where the cost of a missed return is unacceptable.

Key responsibilities:

- Investigate evasion clusters end to end , from a single appeal or signal anomaly to the full linked actor network

- Convert individual findings into durable systemic controls and detection proposals

- Operationalize re-registration controls for high-severity ban populations

- Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities

- Build the recidivism measurement framework: how often banned actors return, how fast we catch them, and which controls reduce return rates

- Author playbooks for contractor-supported evasion review with QA against your own gold standard

- Keep up to date with emerging AI policy enforcement best practices, and use these to inform our decision-making and workflows

Minimum qualifications:

- Experience investigating ban evasion, multi-accounting, or repeat fraud actors at a platform with adversarial users

- Fluency in SQL and comfort building your own analyses across large account and event datasets

- Experience working with fraud or identity-linking signals and a working understanding of their precision/recall tradeoffs

- Rigor about evidence standards , comfort with the asymmetric cost of false positives in severe-harm enforcement

- A track record of turning one-off investigations into repeatable detection logic and policy

- Strong written communication skills, with experience producing clear briefs and recommendations for technical and non-technical stakeholders

- Excellent judgment and the ability to collaborate with team members while navigating rapidly evolving priorities and workstreams

Preferred qualifications:

- Experience using payment or network risk signals in an enforcement context

- Experience with child-safety or other high-severity integrity enforcement

- Experience collaborating directly with detection engineering or data science teams on rule deployment

- A deep interest in AI safety and responsible technology development

- Experience writing effective prompts for generative AI systems in a content review or enforcement context

The annual compensation range for this role is $245,000-$285,000 USD.

## Skills

### Required
- SQL
- fraud investigation
- identity linking
- data analysis
- communication

### Nice to have
- payment risk signals
- child-safety enforcement
- detection engineering
- AI safety
- generative AI systems

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/anthropic/jobs/5319592008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
