Product Manager, Safeguards (Child Safety)
Anthropic · San Francisco, CA · Product Management, Support, & Operations · listed July 7, 2026
The shape of it
Seniority
Manager
Where
Hybrid
Stated pay
$305,000 – $385,000 USD
Requirements listed
9
Length
1,176 words
In the posting’s own words
We are looking for a product manager who is deeply committed to making AI safe and beneficial for humanity. You are aware of the risks and are committed to working with experts and coming up with ideas for Anthropic to implement. You have deep technical expertise in development, deployment and measurement of Safeguards systems. You thrive in rapidly moving and ambiguous environments.
What it asks for · 9
- Ability to make technical tradeoff decisions; ideally with experience working across policy experts, AI/ML research engineers and software engineering teams to design and build state of the art safety systems.
- Strong user understanding of how our products are used, their Safeguards concerns and how we provide the best solutions.
- Demonstrated ability to build product and engineering strategy across multiple cross-functional teams for a rapidly changing space.
- Demonstrated experience in designing and building metrics to evaluate risks, system performance, user impact and making crisp tradeoffs
- Very strong ability to navigate, and prioritize amidst rapidly changing product specs, and to flex into different domains to bring clarity and execute.
- Evidence of exercising judgment and decision making in ambiguous situations.
- Planning, building, launching and measuring new products / systems in a zero to one environment.
- Ability to clearly articulate complex technical concepts to non-technical audiences in written and verbal communication.
- Think creatively about the risks and benefits of new technologies, and think beyond past checklists and playbooks.
Also a plus
- 5+ years in product management with a focus on fast problem understanding, building roadmaps with tractable progress, ability to get into the details on data, detection & interventions, infrastructure & tools, and/or evals.
What the job covers
- Determine how to build in safety by design upstream and leverage downstream defenses for Anthropic’s frontier models, AI products, customers on different surfaces - Claude.ai, 1P API, external Cloud providers.
- Ability to write safety evals and communicate externally about safety.
- Drive impact via ruthless prioritization by clearly defining problems, solution options forward, clarity on both business & technical tradeoffs and accordingly clear requirements toward MVP vs. ideal state.
- Align & collaborate with policy, enforcement, research, engineering and cross functional stakeholders.
- Understand the AI landscape and ecosystem to plan for mitigation of deployment risks of increasingly powerful models and determined adversaries.
- Lead the development of metrics to understand the area, performance, blindspots to help inform future project planning.
Tools and skills named
Models & research
- Evaluations3×
- Machine learning
Product & design
- Product management3×
- User experience
Ways of working
- Cross-functional2×
Words the posting leans on
- product12×
- risks7×
- safeguards6×
- systems5×
- technical5×
- user5×
- build4×
- deployment4×
- research4×
- safety4×
- building3×
- design3×
- development3×
- engineering3×
- evals3×
- tradeoffs3×
Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.
The posting, your resume, and the gaps between them. One click loads all three.
More open at Anthropic
- Account Executive, AI NativeNew York City, NY; San Francisco, CA | New York City, NY
- Account Executive - DNBSingapore
- Account Executive, Public SectorSydney, Australia
- Account Executive - Public Sector (ASEAN)Singapore
- Account Executive, StartupsSan Francisco, CA | New York City, NY
- Accounting, Revenue Internal ControlsSan Francisco, CA | Seattle, WA