Safeguards Enforcement Analyst, Ban Evasion & Recidivism

Anthropic · Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC · Safeguards (Trust & Safety) · listed July 14, 2026

The shape of it

Seniority
Not stated
Where
Hybrid
Stated pay
$245,000 – $285,000 USD
Requirements listed
7
Length
1,049 words

In the posting’s own words

As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be recidivism: a ban that an actor can evade in five minutes isn't enforcement — it's friction. You'll own detecting when banned actors return, linking accounts across identities, and closing the re-registration paths that matter most. The mandate includes our highest-stakes populations, including preventing evasion of child-safety enforcement bans, where the cost of a missed return is unacceptable.

What it asks for · 7

  • Experience investigating ban evasion, multi-accounting, or repeat fraud actors at a platform with adversarial users
  • Fluency in SQL and comfort building your own analyses across large account and event datasets
  • Experience working with fraud or identity-linking signals and a working understanding of their precision/recall tradeoffs
  • Rigor about evidence standards — comfort with the asymmetric cost of false positives in severe-harm enforcement
  • A track record of turning one-off investigations into repeatable detection logic and policy
  • Strong written communication skills, with experience producing clear briefs and recommendations for technical and non-technical stakeholders
  • Excellent judgment and the ability to collaborate with team members while navigating rapidly evolving priorities and workstreams

Also a plus

  • Experience using payment or network risk signals in an enforcement context
  • Experience with child-safety or other high-severity integrity enforcement
  • Experience collaborating directly with detection engineering or data science teams on rule deployment
  • A deep interest in AI safety and responsible technology development
  • Experience writing effective prompts for generative AI systems in a content review or enforcement context

What the job covers

  • Investigate evasion clusters end to end — from a single appeal or signal anomaly to the full linked actor network
  • Convert individual findings into durable systemic controls and detection proposals
  • Operationalize re-registration controls for high-severity ban populations
  • Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities
  • Build the recidivism measurement framework: how often banned actors return, how fast we catch them, and which controls reduce return rates
  • Author playbooks for contractor-supported evasion review with QA against your own gold standard
  • Keep up to date with emerging AI policy enforcement best practices, and use these to inform our decision-making and workflows

Tools and skills named

Languages
  • SQL

Words the posting leans on

  • enforcement11×
  • experience7×
  • actors6×
  • ban4×
  • evasion4×
  • return4×
  • signals4×
  • account3×
  • build3×
  • controls3×
  • detection3×
  • actors return2×
  • banned actors2×
  • child-safety2×
  • comfort2×
  • cost2×

Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.

The posting, your resume, and the gaps between them. One click loads all three.

More open at Anthropic

every open role at Anthropic

How this page was made

An automated read of a public job posting, fetched August 25, 2026 and last changed by Anthropic on August 21, 2026. Every list above is pulled from the posting’s own sentences — nothing rewritten, nothing added, no judgment about the role or the company. Counts and seniority are read off the text by rule, so they can be wrong where the posting is unusual. The original is the only thing that binds. Openings close without warning; check the source before spending an evening on it.