Review of the CB risk determination in the Claude Mythos 5.1 System Card

·LessWrong··

This post is best viewed on the MCNAIR website. Before the release of their latest publicly-known model, Claude Mythos 5.1, Anthropic conducted several human-run and automated evaluations to assess the risk of their models allowing a well-resourced team to replace the extremely specialized expertise needed to design and deploy a novel chemical or biological weapon (the CB-2 threshold)[1]. They concluded that Mythos was unable to do so.Independent assessment of the public evidence:Overall, we agr...

Read full article →

Related Articles

Formalizing Fermat's Last Theorem
jlebar · Hacker News · 1d ago
Why are European countries moving their gold out of North America?
ranit · Hacker News · 19h ago
Falsehoods Programmers Believe About LANs
robinpie · Hacker News · 3h ago
Hackers Had a Live Feed of Every ID Verification Company Scanned for over a Year
beardyw · Hacker News · 1d ago
Artificial Analysis Intelligence Index v4.2
nojs · Hacker News · 1d ago