InternFlow

Greenhouse ·

full-time

Product Manager, Safeguards (account Integrity & Abuse)

Anthropic · San Francisco, CA | New York City, NY

Education

  • Level: Bachelor's degree
Product Management

Key Skills

Product ManagementAI SafetyMetricsTechnical Communication

About the Role Anthropic’s Safeguards team protects users from risks of powerful AIs. The Product Manager will own ideation, design, development and deployment of Safeguards systems and product UX across cloud platforms, working with research and product teams to develop detections, evaluations, interventions, and measurement tools. Responsibilities Determine how to build safety by design upstream and leverage downstream defenses for Anthropic’s frontier models and products. Write safety evaluations and communicate externally about safety. Prioritize problems, define solution options, and set clear requirements for MVP versus ideal state. Align and collaborate with policy, enforcement, research, engineering and cross‑functional stakeholders. Understand AI landscape to plan mitigation of deployment risks. Lead development of metrics to assess performance, blind spots and inform project planning. Requirements Ability to make technical trade‑off decisions and work with policy experts, AI/ML researchers and software engineers. Strong user understanding of product usage and Safeguards concerns. Experience designing and building metrics to evaluate risks, system performance and user impact. Ability to navigate rapidly changing product specs, exercise judgment in ambiguous situations, and launch zero‑to‑one products. Excellent communication of complex technical concepts to non‑technical audiences. Preferred: 5+ years product management experience focusing on data, detection, interventions, infrastructure and tools.

Skills for this role

Product ManagementAI SafetyMetricsTechnical Communication
RoleProduct Manager, Safeguards (account Integrity & Abuse)
CompanyAnthropic
LocationSan Francisco, CA | New York City, NY
CompensationNot disclosed
Posted2026-09-21
DeadlineRolling

Typical process for this type of role

A general guide — the exact steps for this specific listing may vary; check the original posting for details.

  1. 1ApplicationSubmit your resume through the apply link.
  2. 2ScreeningRecruiter reviews your background against the role.
  3. 3AssessmentA technical test, assignment, or coding round, depending on the role.
  4. 4Interview(s)One or more rounds with the hiring team.
  5. 5OfferOffer letter with compensation and start date.

Before you apply

0/4

Product Manager, Safeguards (account Integrity & Abuse) at Anthropic: frequently asked questions

Who can apply for the Product Manager, Safeguards (account Integrity & Abuse) role at Anthropic?
The listing asks for Bachelor's degree; freshers are welcome.
What skills does the Product Manager, Safeguards (account Integrity & Abuse) role require?
The listing highlights Product Management, AI Safety, Metrics, Technical Communication. Show each of these in a project or past role on your resume.
What is the salary for this role?
Anthropic has not stated compensation in the listing. Check the original posting or ask during the application process.
Is the Product Manager, Safeguards (account Integrity & Abuse) position remote, hybrid or onsite?
The listing gives San Francisco, CA | New York City, NY as the location and does not state a work mode.
What is the application deadline?
Anthropic has not listed a fixed deadline, so apply early in case the opening is filled.
How do I apply for the Product Manager, Safeguards (account Integrity & Abuse) role?
Use the Apply button on this page. It opens the original listing on job-boards.greenhouse.io, where you submit your application with the company.

More at Anthropic

Other jobs at Anthropic

See all 122 openings at Anthropic

Explore Related Placements


// similar opportunities

You might also like

Product Manager, Safe Access

Anthropic

San Francisco, CA | New York City, NYfull-time
Bachelor'sGreenhouseProduct Management# Product Management# API Design# User Experience (UX)# Identity Management+2 more

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...

Apply now →

Program Manager, Brand

Anthropic

San Francisco, CA | New York City, NYfull-time
HybridBachelor'sGreenhouseProgram Management# Program Management# Brand Marketing# Project Management# Strategic Planning+5 more

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...

['Flexibility to travel', 'Hybrid work policy', 'Annual salary $200k-$255k']

Apply now →

Senior Product Manager, Integrity & Trust Experience

Cloudflare

Hybridfull-time
Greenhouse

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of ...

Apply now →

Technical Program Manager, Life Sciences

Anthropic

San Francisco, CA | New York City, NYfull-time
Greenhouse

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...

Apply now →

Commercial Operations Program Manager

Anthropic

San Francisco, CA | New York City, NYfull-time
Greenhouse

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...

Apply now →

Staff+ Software Engineer, Account Abuse (machine Learning)

Anthropic

San Francisco, CA | New York City, NYfull-time
Greenhouse

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for s...

Apply now →