Technical Program Manager, Safeguards (Infrastructure & Evals)

Anthropic · San Francisco, CA | New York City, NY | Seattle, WA · Technical Program Management · listed February 4, 2026

The shape of it

Seniority
Manager
Where
Not stated
Stated pay
$290,000 – $365,000 USD
Requirements listed
2
Length
1,523 words

In the posting’s own words

As a Technical Program Manager for Safeguards Infrastructure and Evals, you'll own the operational health and forward momentum of this stack. Your primary responsibility is driving reliability — owning the incident-response and post-mortem process, ensuring SLOs are defined and met in partnership with various teams, and making sure that when things go wrong, the right people know, the right actions get taken, and those actions actually get closed out. Alongside that ongoing operational rhythm, you'll coordinate the larger platform investments: migrations, eval-platform improvements, and the cross-team dependencies that connect them.

What it asks for · 2

  • Have experience with or strong interest in AI safety — you understand why the reliability of a safety-critical pipeline is a different kind of problem than the reliability of a product feature, and that distinction motivates you.
  • Have experience with SRE practices, incident management frameworks, or on-call operations at scale.

Also a plus

  • Have experience with SRE practices, incident management frameworks, or on-call operations at scale.
  • Have worked on or with evaluation infrastructure for ML systems — understanding how evals get designed, run, and interpreted.
  • Have experience driving infrastructure migrations in complex, multi-team environments — particularly where the migration touches operational systems that can't go offline.
  • Be familiar with monitoring and alerting tooling (PagerDuty, Datadog, or equivalents) and the operational culture around them.

What the job covers

  • Own the Safeguards Engineering ops review - Drive the recurring cadence that keeps the team informed and coordinated: surfacing recent incidents and failures, bringing visibility to reliability trends, and making sure the right people are in the room when decisions need to be made. This is the heartbeat of how Safeguards Eng stays ahead of operational risk.
  • Drive incident tracking and post-mortem execution - When incidents happen — and in this space, they happen regularly — you'll make sure they get followed through properly. That means tracking incidents across the organization (including those owned by partner teams like Inference), ensuring post-mortems get written, and most critically, making sure the action items that come out of them actually get done. Closing the loop on post-mortem actions is one of the highest-leverage things this role does.
  • Establish and maintain SLOs with partner teams - Work with Safeguards Engineering teams and key partners — particularly Inference and Cloud Inference — to define service-level objectives for safety-critical pipelines. Then build the tracking and reporting that makes it possible to tell whether those SLOs are being met, and surface it when they're not.
  • Maintain runbook quality and incident-ownership clarity - Safety-critical systems need clear playbooks for when things go wrong. Partner with engineering leads to keep runbooks accurate, actionable, and up to date — and ensure that ownership of incidents is unambiguous so that nothing falls through the cracks during an active incident.
  • Drive platform migrations and infrastructure projects - Own the program management for the larger infrastructure work on the roadmap: migrating the infra from one platform to the next, moving from one incident platform to the next and from one cloud system monitoring to another, and other migrations as they come. These are cross-team efforts with real dependencies — your job is to keep them sequenced, on track, and connected to the teams that need them.
  • Coordinate evals platform improvements - Partner with the evals engineering team to drive improvements to the evaluation platform — including self-serve capabilities and the broader eval factory infrastructure. Help scope the work, track dependencies on other Safeguards systems, and make sure the evals platform is keeping pace with the team's needs.

Tools and skills named

Models & research
  • Evaluations5×
  • Inference4×
  • Machine learning2×
Operations & finance
  • Program management4×
Cloud & infra
  • Datadog
  • Site reliability
Product & design
  • Roadmap
Ways of working
  • On-call

Words the posting leans on

  • incident10×
  • platform9×
  • systems9×
  • infrastructure8×
  • get7×
  • operational7×
  • evals6×
  • partner6×
  • safeguards6×
  • actions5×
  • keep5×
  • migrations5×
  • post-mortem5×
  • safety-critical5×
  • sure5×
  • build4×

Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.

The posting, your resume, and the gaps between them. One click loads all three.

More open at Anthropic

every open role at Anthropic

How this page was made

An automated read of a public job posting, fetched August 24, 2026 and last changed by Anthropic on August 21, 2026. Every list above is pulled from the posting’s own sentences — nothing rewritten, nothing added, no judgment about the role or the company. Counts and seniority are read off the text by rule, so they can be wrong where the posting is unusual. The original is the only thing that binds. Openings close without warning; check the source before spending an evening on it.