Safeguards Enforcement Analyst, Cyber Harm
Anthropic · Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC
Posted about 2 months ago
or apply directly on Anthropic's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 66 days Anthropic's roles stay open a median of 62 days
- Reposted
- No
- Salary listed
- No 0% of Anthropic's roles list one
- Ghost-job risk at Anthropic
- high 449 stale, 19 reposted of 603 open
- Hiring momentum
- 795 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-10-08
Measured from postings appearing on and disappearing from Anthropic's own greenhouse board since 2026-08-03. Full hiring picture for Anthropic.
About this role
As a Safeguards Enforcement Analyst at Anthropic, you will review flagged content and execute enforcement actions to prevent the misuse of AI systems, particularly in the context of cyberattacks and malware. Your role involves detecting potential threats, providing feedback on policy gaps, and collaborating with engineering and data science teams to improve detection models. You may also be required to respond to escalations during weekends and holidays, and you will engage with sensitive content as part of your responsibilities.
- benefits
- 3/5
- freshness
- 1/5
- career value
- 4/5
- role clarity
- 4/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Anthropic as an employer.
What you need
- Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research
- Experience performing content review, abuse investigations, or policy enforcement at volume
- Proficiency in SQL and/or Python for data analysis and threat detection
- Experience identifying emerging risks and communicating findings to diverse stakeholders, such as Product, Policy, Engineering, and Legal teams
- Experience working with generative AI products, including writing effective prompts for content review and enforcement
Nice to have
- Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company
- Experience with large language models and an understanding of how AI technology could be misused for cyber operations
- Experience operating within abuse monitoring programs or enforcement review systems
- Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
- Experience working with government agencies, regulated environments, or information sharing communities
What you get
- Annual Salary: $285,000 — $330,000 USD
- Competitive compensation and benefits
- Optional equity donation matching
- Generous vacation and parental leave
- Flexible working hours
- Lovely office space for collaboration
Worth weighing
- Exposure to explicit content spanning violent, technical, or psychologically disturbing topics
- Role may require responding to escalations during weekends and holidays
- Visa sponsorship available but not guaranteed for every role and candidate
Summarised from Anthropic's posting. Read the full original.
Listed by Anthropic on their greenhouse job board, last confirmed open on 2026-10-08. PitchMeAI is not the employer.
More roles at Anthropic
- Staff Software Engineer, Environments InfrastructureSan Francisco, CA | New York City, NY
- Research Engineer, Cybersecurity RL (Reinforcement Learning)San Francisco, CA | New York City, NY
- Technical Program Manager, API PlatformSan Francisco, CA | Seattle, WA
- Senior Manager, Technical Accounting - M&A and InvestmentsSan Francisco, CA | Seattle, WA
- Product Operations Manager, EmbeddedSan Francisco, CA | New York City, NY | Seattle, WA
- Staff Engineer, Datacenter Server LifecycleSydney, Australia
- Staff+ Software Engineer, Capacity EngineeringSan Francisco, CA | New York City, NY | Seattle, WA
- Applied AI Architect, PartnershipsSan Francisco, CA | New York City, NY