Greenhouse ·
full-timeProduct Manager, Safeguards (account Integrity & Abuse)
Anthropic · San Francisco, CA | New York City, NY
Education
- Level: Bachelor's degree
Key Skills
About the Role Anthropic’s Safeguards team protects users from risks of powerful AIs. The Product Manager will own ideation, design, development and deployment of Safeguards systems and product UX across cloud platforms, working with research and product teams to develop detections, evaluations, interventions, and measurement tools. Responsibilities Determine how to build safety by design upstream and leverage downstream defenses for Anthropic’s frontier models and products. Write safety evaluations and communicate externally about safety. Prioritize problems, define solution options, and set clear requirements for MVP versus ideal state. Align and collaborate with policy, enforcement, research, engineering and cross‑functional stakeholders. Understand AI landscape to plan mitigation of deployment risks. Lead development of metrics to assess performance, blind spots and inform project planning. Requirements Ability to make technical trade‑off decisions and work with policy experts, AI/ML researchers and software engineers. Strong user understanding of product usage and Safeguards concerns. Experience designing and building metrics to evaluate risks, system performance and user impact. Ability to navigate rapidly changing product specs, exercise judgment in ambiguous situations, and launch zero‑to‑one products. Excellent communication of complex technical concepts to non‑technical audiences. Preferred: 5+ years product management experience focusing on data, detection, interventions, infrastructure and tools.
Skills for this role
| Role | Product Manager, Safeguards (account Integrity & Abuse) |
|---|---|
| Company | Anthropic |
| Location | San Francisco, CA | New York City, NY |
| Compensation | Not disclosed |
| Posted | 2026-09-21 |
| Deadline | Rolling |
Typical process for this type of role
A general guide — the exact steps for this specific listing may vary; check the original posting for details.
- 1ApplicationSubmit your resume through the apply link.
- 2ScreeningRecruiter reviews your background against the role.
- 3AssessmentA technical test, assignment, or coding round, depending on the role.
- 4Interview(s)One or more rounds with the hiring team.
- 5OfferOffer letter with compensation and start date.
Before you apply
0/4Product Manager, Safeguards (account Integrity & Abuse) at Anthropic: frequently asked questions
- Who can apply for the Product Manager, Safeguards (account Integrity & Abuse) role at Anthropic?
- The listing asks for Bachelor's degree; freshers are welcome.
- What skills does the Product Manager, Safeguards (account Integrity & Abuse) role require?
- The listing highlights Product Management, AI Safety, Metrics, Technical Communication. Show each of these in a project or past role on your resume.
- What is the salary for this role?
- Anthropic has not stated compensation in the listing. Check the original posting or ask during the application process.
- Is the Product Manager, Safeguards (account Integrity & Abuse) position remote, hybrid or onsite?
- The listing gives San Francisco, CA | New York City, NY as the location and does not state a work mode.
- What is the application deadline?
- Anthropic has not listed a fixed deadline, so apply early in case the opening is filled.
- How do I apply for the Product Manager, Safeguards (account Integrity & Abuse) role?
- Use the Apply button on this page. It opens the original listing on job-boards.greenhouse.io, where you submit your application with the company.
More at Anthropic
Other jobs at Anthropic
- Staff + Software Security Engineer, Labs · San Francisco, CA | New York City, NY | Seattle, WA
- Data Center Engineer, Reliability & Infrastructure Management – Compute Supply · San Francisco, CA | New York City, NY
- Security Risk & Compliance, Agent Security · San Francisco, CA | New York City, NY
- Senior Manager, M&A Finance Integration · San Francisco, CA | Seattle, WA
- Manager, Applied AI Engineering (megas) · San Francisco, CA | New York City, NY; San Francisco, CA | Seattle, WA
Explore Related Placements
// similar opportunities
You might also like
Product Manager, Safe Access
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...
Program Manager, Brand
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...
['Flexibility to travel', 'Hybrid work policy', 'Annual salary $200k-$255k']
Senior Product Manager, Integrity & Trust Experience
Cloudflare
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of ...
Technical Program Manager, Life Sciences
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...
Commercial Operations Program Manager
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...
Staff+ Software Engineer, Account Abuse (machine Learning)
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...