Product Manager, Safeguards (Child Safety)

Anthropic · San Francisco, CA · Product Management, Support, & Operations · listed July 7, 2026

The shape of it

Seniority
Manager
Where
Hybrid
Stated pay
$305,000 – $385,000 USD
Requirements listed
9
Length
1,176 words

In the posting’s own words

We are looking for a product manager who is deeply committed to making AI safe and beneficial for humanity. You are aware of the risks and are committed to working with experts and coming up with ideas for Anthropic to implement. You have deep technical expertise in development, deployment and measurement of Safeguards systems. You thrive in rapidly moving and ambiguous environments.

What it asks for · 9

  • Ability to make technical tradeoff decisions; ideally with experience working across policy experts, AI/ML research engineers and software engineering teams to design and build state of the art safety systems.
  • Strong user understanding of how our products are used, their Safeguards concerns and how we provide the best solutions.
  • Demonstrated ability to build product and engineering strategy across multiple cross-functional teams for a rapidly changing space.
  • Demonstrated experience in designing and building metrics to evaluate risks, system performance, user impact and making crisp tradeoffs
  • Very strong ability to navigate, and prioritize amidst rapidly changing product specs, and to flex into different domains to bring clarity and execute.
  • Evidence of exercising judgment and decision making in ambiguous situations.
  • Planning, building, launching and measuring new products / systems in a zero to one environment.
  • Ability to clearly articulate complex technical concepts to non-technical audiences in written and verbal communication.
  • Think creatively about the risks and benefits of new technologies, and think beyond past checklists and playbooks.

Also a plus

  • 5+ years in product management with a focus on fast problem understanding, building roadmaps with tractable progress, ability to get into the details on data, detection & interventions, infrastructure & tools, and/or evals.

What the job covers

  • Determine how to build in safety by design upstream and leverage downstream defenses for Anthropic’s frontier models, AI products, customers on different surfaces - Claude.ai, 1P API, external Cloud providers.
  • Ability to write safety evals and communicate externally about safety.
  • Drive impact via ruthless prioritization by clearly defining problems, solution options forward, clarity on both business & technical tradeoffs and accordingly clear requirements toward MVP vs. ideal state.
  • Align & collaborate with policy, enforcement, research, engineering and cross functional stakeholders.
  • Understand the AI landscape and ecosystem to plan for mitigation of deployment risks of increasingly powerful models and determined adversaries.
  • Lead the development of metrics to understand the area, performance, blindspots to help inform future project planning.

Tools and skills named

Models & research
  • Evaluations3×
  • Machine learning
Product & design
  • Product management3×
  • User experience
Ways of working
  • Cross-functional2×

Words the posting leans on

  • product12×
  • risks7×
  • safeguards6×
  • systems5×
  • technical5×
  • user5×
  • build4×
  • deployment4×
  • research4×
  • safety4×
  • building3×
  • design3×
  • development3×
  • engineering3×
  • evals3×
  • tradeoffs3×

Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.

The posting, your resume, and the gaps between them. One click loads all three.

More open at Anthropic

every open role at Anthropic

How this page was made

An automated read of a public job posting, fetched August 24, 2026 and last changed by Anthropic on August 21, 2026. Every list above is pulled from the posting’s own sentences — nothing rewritten, nothing added, no judgment about the role or the company. Counts and seniority are read off the text by rule, so they can be wrong where the posting is unusual. The original is the only thing that binds. Openings close without warning; check the source before spending an evening on it.