Safeguards Enforcement Analyst, Integrity & Authenticity
Anthropic · Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC · Safeguards (Trust & Safety) · listed July 10, 2026
The shape of it
Seniority
Not stated
Where
Hybrid
Stated pay
$285,000 – $330,000 USD
Requirements listed
6
Length
1,203 words
In the posting’s own words
As a Safeguards Analyst focusing on Integrity & Authenticity, you will be responsible for building and executing enforcement workflows for our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for coordinated inauthentic behavior, election manipulation, and targeting, tracking, and surveillance of individuals.
What it asks for · 6
- Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field with a focus on one or more of: influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms
- Experience standing up and scaling policy enforcement or content review workflows
- Proficiency in SQL and/or other data analysis tools to draw insights from large datasets
- Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams
- Experience working with generative AI products, including writing effective prompts for content review and enforcement
- Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
Also a plus
- Experience conducting cross-platform investigations into influence operations, coordinated inauthentic behavior, or disinformation campaigns
- Familiarity with open-source intelligence (OSINT) techniques and tools used for threat actor tracking and network analysis
- Working knowledge of privacy law, surveillance technology, or data broker ecosystems as they relate to targeting and tracking harms
- Experience with large language models and an understanding of how AI technology could be misused to generate synthetic personas, fabricate quotes, or automate persuasion at scale
- Familiarity with election security frameworks, campaign finance law, or electoral integrity standards in one or more jurisdictions
- Experience navigating evolving regulatory landscapes relevant to this space (e.g., DSA, EU AI Act, FEC regulations, GDPR)
- Experience working with election bodies, civil society organizations, or government agencies on integrity or disinformation-related issues
- Proficiency in Python for data analysis and automation
What the job covers
- Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy
- Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement systems
- Review flagged content to drive enforcement and policy improvements
- Enforce usage policies with a focus on detecting and mitigating AI-enabled influence operations, coordinated inauthentic behavior, election interference, and targeting, tracking, or surveillance of individuals and groups
- Support the Safeguards policy design team by providing detailed feedback on policy gaps based on real enforcement scenarios
- Keep up to date with emerging AI policy enforcement best practices, evolving threat actor tactics, and the regulatory landscape around elections, privacy, and surveillance, using these to inform our decision-making and workflows
Tools and skills named
Security & compliance
- Regulatory2×
- GDPR
- Security
Languages
- Python
- SQL
Models & research
- LLM
Words the posting leans on
- enforcement10×
- experience9×
- policy9×
- election6×
- surveillance6×
- content5×
- products5×
- threat5×
- tracking5×
- actor4×
- coordinated inauthentic4×
- data4×
- inauthentic behavior4×
- influence operations4×
- integrity4×
- policy enforcement4×
Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.
The posting, your resume, and the gaps between them. One click loads all three.
More open at Anthropic
- Account Executive, AI NativeNew York City, NY; San Francisco, CA | New York City, NY
- Account Executive - DNBSingapore
- Account Executive, Public SectorSydney, Australia
- Account Executive - Public Sector (ASEAN)Singapore
- Account Executive, StartupsSan Francisco, CA | New York City, NY
- Accounting, Revenue Internal ControlsSan Francisco, CA | Seattle, WA