or apply directly on Anthropic's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 15 days Anthropic's roles stay open a median of 40 days
- Reposted
- No
- Salary listed
- No 0% of Anthropic's roles list one
- Ghost-job risk at Anthropic
- high 276 stale, 19 reposted of 603 open
- Hiring momentum
- 795 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-17
Measured from postings appearing on and disappearing from Anthropic's own greenhouse board since 2026-08-03. Full hiring picture for Anthropic.
About this role
As a Cyber Evaluations Engineer at Anthropic, you will be responsible for designing and conducting evaluations to assess the cyber-relevant capabilities and robustness of AI models. Your day-to-day tasks will include running safeguard-robustness testing, analyzing evaluation results, and collaborating with cross-functional teams to enhance detection architectures against cyber misuse.
- benefits
- 3/5
- freshness
- 4/5
- career value
- 4/5
- role clarity
- 4/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Anthropic as an employer.
What you need
- Experience building or running evaluations, benchmarks, or test suites for software or ML systems, including delivering results on short, fixed timelines
- Hands-on cybersecurity experience (e.g., CTF participation, vulnerability research, exploit development, or security research)
- Proficiency in Python
- Strong ability to communicate evaluation results with multiple cross-functional stakeholders or potential policy stakeholders
Nice to have
- Deep offensive-security or security-research experience, including experience building AI security benchmarks
- Experience analyzing adversarial or abuse data (e.g., jailbreaks, prompt bypasses, intrusion or fraud telemetry)
- Experience working onsite with government partners on testing or evaluation engagements
- Experience with AI/ML evaluation frameworks
- Familiarity with coordinated vulnerability disclosure practices
What you get
- Competitive compensation and benefits
- Optional equity donation matching
- Generous vacation and parental leave
- Flexible working hours
- Lovely office space for collaboration
Worth weighing
- No specific mention of the technology stack used in evaluations
- Role involves significant collaboration with policy teams, which may require navigating complex regulatory environments
- Visa sponsorship is available but not guaranteed for every role or candidate
- Location-based hybrid policy requires at least 25% in-office presence
Summarised from Anthropic's posting. Read the full original.
Listed by Anthropic on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.
More roles at Anthropic
- Enterprise Integrated Campaign ManagerSan Francisco, CA | New York City, NY | Seattle, WA
- Finance Systems Engineer, Finance and StrategySan Francisco, CA
- Strategic Account Executive, IndustriesSan Francisco, CA | New York City, NY
- Strategy & Operations Lead, Enterprise MarketingSan Francisco, CA | New York City, NY
- Customer Marketing Manager, IndustriesSan Francisco, CA | New York City, NY
- Staff+ Software Engineer, Platform EcosystemSan Francisco, CA | New York City, NY
- Software Engineering Manager, Network SecuritySan Francisco, CA | New York City, NY
- Head of Treasury Strategy & TransformationRemote-Friendly (Travel Required) | San Francisco, CA