Staff+ Software Engineer, Safeguards
Anthropic · San Francisco, CA | New York City, NY
Posted about 2 months ago
or apply directly on Anthropic's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 66 days Anthropic's roles stay open a median of 62 days
- Reposted
- No
- Salary listed
- No 0% of Anthropic's roles list one
- Ghost-job risk at Anthropic
- high 449 stale, 19 reposted of 603 open
- Hiring momentum
- 795 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-10-08
Measured from postings appearing on and disappearing from Anthropic's own greenhouse board since 2026-08-03. Full hiring picture for Anthropic.
About this role
As a Staff+ Software Engineer on the Safeguards team at Anthropic, you will develop systems to monitor AI models, prevent misuse, and ensure user well-being. Your responsibilities will include building monitoring and abuse detection mechanisms, creating internal dashboards for analysts, and collaborating with research teams to enhance model safety. The role emphasizes technical skills in software engineering while upholding safety and transparency principles.
- benefits
- 3/5
- freshness
- 1/5
- career value
- 4/5
- role clarity
- 4/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Anthropic as an employer.
What you need
- Bachelor’s degree in Computer Science, Software Engineering or comparable professional experience
- Proficiency in at least one programming language such as Python, Java, etc.
- Ability to work across the stack
- Strong communication skills and ability to explain complex technical concepts to non-technical stakeholders
Nice to have
- 8+ years of experience in a software engineering position
- Experience with integrity, spam, fraud, or abuse detection and mitigation
- Experience building trust and safety detection mechanisms and intervention for AI/ML systems
- Experience with prompt engineering, jailbreak attacks, and other adversarial inputs
- Experience working closely with operational teams to build custom internal tooling
What you get
- Competitive compensation
- Optional equity donation matching
- Generous vacation and parental leave
- Flexible working hours
- Lovely office space for collaboration
Worth weighing
- No specific team placement mentioned until after the interview process
- Visa sponsorship is available but not guaranteed for every role
- Location-based hybrid policy requires at least 25% in-office presence
Summarised from Anthropic's posting. Read the full original.
Listed by Anthropic on their greenhouse job board, last confirmed open on 2026-10-08. PitchMeAI is not the employer.
More roles at Anthropic
- Staff Software Engineer, Environments InfrastructureSan Francisco, CA | New York City, NY
- Research Engineer, Cybersecurity RL (Reinforcement Learning)San Francisco, CA | New York City, NY
- Technical Program Manager, API PlatformSan Francisco, CA | Seattle, WA
- Senior Manager, Technical Accounting - M&A and InvestmentsSan Francisco, CA | Seattle, WA
- Product Operations Manager, EmbeddedSan Francisco, CA | New York City, NY | Seattle, WA
- Staff Engineer, Datacenter Server LifecycleSydney, Australia
- Staff+ Software Engineer, Capacity EngineeringSan Francisco, CA | New York City, NY | Seattle, WA
- Applied AI Architect, PartnershipsSan Francisco, CA | New York City, NY