Landmark new METR report: Can AIs already start ‘rogue deployments’ inside AI companies? by 80000_Hours

·Nuno Sempere··

By Robert Wiblin | Watch on Youtube | Listen on SpotifyA red-teamer was em­bed­ded in­side An­thropic for three weeks, told to imag­ine he was an evil Claude, and asked to figure out how to launch a ‘rogue AI de­ploy­ment’ with­out get­ting caught.It’s one part of a land­mark new re­port from METR — the out­fit be­hind the task-com­ple­tion time hori­zon graph which has be­come the sin­gle most watched mea­sure of AI progress.This ma­jor new re­search push is be­ing con­ducted with close col­lab...

Read full article →

Related Articles

What is recursive self-improvement, and what would it mean to ban it? by sarahhw
sarahhw · Nuno Sempere · 20h ago
State of the Field: AI for Epistemics and Coordination by Ben_N
Ben_N · Nuno Sempere · 20h ago
Will a non-lab AI-only effort solve a Millennium Problem in 2026?
Bayesian · Manifold Markets · 23h ago
Will the U.S. use a nuclear weapon against Iran’s Pickaxe Mountain between November 4 and November 10, 2026?
Laurent27bis · Manifold Markets · 1d ago
How many people will attend EAG NYC 2026?
d · Manifold Markets · 1d ago