Anthropic

Safeguards Enforcement Analyst, Integrity & Authenticity

Anthropic · Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC

Posted about 2 months ago

or apply directly on Anthropic's site. We never take the application ourselves.

Is this posting real?

This role has been open
66 days
Anthropic's roles stay open a median of 62 days
Reposted
No
Salary listed
No
0% of Anthropic's roles list one
Ghost-job risk at Anthropic
high
449 stale, 19 reposted of 603 open
Hiring momentum
795 roles opened in the last 90 days
↑ up vs. the prior 90 days
Last confirmed on the employer's board
2026-10-08

Measured from postings appearing on and disappearing from Anthropic's own greenhouse board since 2026-08-03. Full hiring picture for Anthropic.

About this role

As a Safeguards Enforcement Analyst at Anthropic, you will design and implement enforcement workflows to detect and mitigate misuse of AI systems, particularly regarding inauthentic behavior and election interference. Your role involves collaborating with engineering and data science teams to optimize detection models, reviewing flagged content, and providing feedback on policy gaps. You will also stay informed about emerging threats and best practices in AI policy enforcement.

Our read on this posting2.4out of 5
benefits
3/5
freshness
1/5
career value
4/5
role clarity
4/5
pay transparency
0/5

Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Anthropic as an employer.

What you need

  • Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field with a focus on influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms
  • Experience standing up and scaling policy enforcement or content review workflows
  • Proficiency in SQL and/or other data analysis tools to draw insights from large datasets
  • Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement
  • Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space

Nice to have

  • Experience conducting cross-platform investigations into influence operations, coordinated inauthentic behavior, or disinformation campaigns
  • Familiarity with open-source intelligence (OSINT) techniques and tools used for threat actor tracking and network analysis
  • Working knowledge of privacy law, surveillance technology, or data broker ecosystems as they relate to targeting and tracking harms
  • Experience with large language models and an understanding of how AI technology could be misused to generate synthetic personas, fabricate quotes, or automate persuasion at scale
  • Familiarity with election security frameworks, campaign finance law, or electoral integrity standards in one or more jurisdictions

What you get

  • Competitive compensation and benefits
  • Optional equity donation matching
  • Generous vacation and parental leave
  • Flexible working hours
  • Lovely office space for collaboration

Worth weighing

  • Role may require responding to escalations during weekends and holidays, particularly around major electoral events
  • Exposure to explicit content spanning political, violent, or psychologically disturbing topics
  • Visa sponsorship available but not guaranteed for every role and candidate

Summarised from Anthropic's posting. Read the full original.

Listed by Anthropic on their greenhouse job board, last confirmed open on 2026-10-08. PitchMeAI is not the employer.

More roles at Anthropic

All 603 open roles at Anthropic →

Apply to a Safeguards Enforcement Analyst, Integrity & Authenticity position at Anthropic - Jobs near me at Remote-Friendly, United States; San Francisco and 3 more locations