Research Manager, Biological Safety

Anthropic · San Francisco, CA · Safeguards (Trust & Safety) · listed September 2, 2026

The shape of it

Seniority
Manager
Experience asked
3–8 years
Where
Not stated
Stated pay
$405,000 – $485,000 USD
Requirements listed
10
Length
1,912 words

In the posting’s own words

You will lead a team of research scientists and engineers who design and run capability evaluations against frontier models, curate training data for our safety classifiers, train and iterate on those classifiers alongside our ML engineers, and measure how they hold up against adversarial pressure in production traffic. You will set the technical direction for that work, decide where the team invests, and own the results.

What it asks for · 10

  • Experience managing a technical team, including hiring, coaching, and performance management
  • A record of setting technical direction for a team and making prioritization calls under uncertainty
  • Proficiency in Python, with a background in scientific programming and data analysis
  • A solid grasp of ML fundamentals, sufficient to critically review evaluation design and classifier development
  • Knowledge of modern biology across both measurement and engineering: high-throughput assays and functional characterization, as well as gene synthesis, genome editing, strain construction, and protein engineering
  • Experience designing quantitative experiments or evaluations and drawing defensible conclusions from noisy results
  • Clear analytical and writing skills, and the ability to explain technical concepts to non-technical stakeholders
  • Familiarity with dual-use research concerns and biosecurity frameworks, such as select agent regulations, the Biological Weapons Convention, or Australia Group guidelines
  • Comfort with ambiguity and with shifting priorities as AI capabilities change
  • Motivation to prevent misuse without obstructing the beneficial work that makes up the vast majority of this field

Also a plus

  • 3+ years of people management experience, ideally leading research scientists, research engineers, or ML engineers
  • Experience building a team or function from a small headcount, including defining scope, hiring the first few people, and establishing how the team works
  • At least 8 years of hands-on experience in life sciences, with deep expertise in areas such as molecular biology, drug discovery, or computational biology
  • Experience working with large language models, including prompting, fine-tuning, or evaluation
  • Experience training or deploying classifiers or other ML systems in production, and comfort reasoning about precision and recall for rare, high-consequence categories where the base rate is very low
  • Experience developing ML methods for biological systems or biological data
  • Familiarity with adversarial robustness, red-teaming, or safety evaluation of ML systems
  • Experience leading complex technical projects across multiple stakeholder groups

What the job covers

  • Manage, coach, and grow a team of research scientists and engineers working on biological safety evaluations and classifiers, including hiring, onboarding, performance, and career development
  • Set the technical direction and roadmap for the biological safety research agenda, and make the calls about what the team builds, what it deprioritizes, and when a safeguard is ready to ship
  • Own the quality of capability evaluations that assess what new models can do in the biological domain, and turn results into deployment recommendations that leadership can act on
  • Guide the development of training and evaluation datasets for our safety classifiers, working with internal and external threat modeling experts to ground them in realistic risk
  • Oversee the training and iteration of safety classifiers alongside ML engineers, optimizing jointly for adversarial robustness and low false-positive rates
  • Ensure the team invests in the tooling and pipelines that make evaluation and classifier development fast and repeatable
  • Establish how the team measures classifier and eval performance against production traffic, identifies gaps, and prioritizes improvements
  • Direct red-teaming and stress-testing of safeguards as threats, models, and product surfaces evolve
  • Partner with Research, Product, Policy, and government affairs colleagues to embed biological safety throughout the model development lifecycle, and serve as an escalation point for biological content
  • Represent the team's work in external communications including model cards, blog posts, and policy documents
  • Track developments in biology, machine learning, and biosecurity for their potential to create new risks or enable new mitigations

Tools and skills named

Models & research
  • Machine learning15×
  • Evaluations9×
  • Fine-tuning2×
  • LLM2×
  • Prompt engineering2×
Ways of working
  • Mentorship2×
  • Testing2×
Languages
  • Python2×
Product & design
  • Roadmap2×
Security & compliance
  • Threat modeling2×

Words the posting leans on

  • evaluation19×
  • biological18×
  • classifiers18×
  • experience18×
  • research15×
  • safety14×
  • models13×
  • development12×
  • technical12×
  • engineers10×
  • biology8×
  • biological safety7×
  • systems7×
  • training7×
  • performance6×
  • safeguards6×

Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.

The posting, your resume, and the gaps between them. One click loads all three.

More open at Anthropic

every open role at Anthropic

How this page was made

An automated read of a public job posting, fetched September 2, 2026 and last changed by Anthropic on September 2, 2026. Every list above is pulled from the posting’s own sentences — nothing rewritten, nothing added, no judgment about the role or the company. Counts and seniority are read off the text by rule, so they can be wrong where the posting is unusual. The original is the only thing that binds. Openings close without warning; check the source before spending an evening on it.