Rerunning AI safety papers on every frontier release would be pretty easy and valuable

·LessWrong··

tl;dr: Some important AI safety research is never rerun on the newest models. There are probably cases where this would be valuable and a single well-positioned researcher could likely do this with sufficient funding.This summer, Second Look Research (SLR) is running a summer fellowship dedicated to empirical replications of AI safety research. Many of our most interesting results so far came from replicating previous results on newer or more capable models. For example, it is perhaps useful to ...

Read full article →

Related Articles

U.S. Strategic Petroleum Reserve Falls to Lowest Level Since 1982
thelastgallon · Hacker News · 7h ago
Does Reddit have an astroturfing problem? What the data suggests
p-s-v · Hacker News · 20h ago
Nvidia wants to put a watchdog chip next to every AI agent
jonbaer · Hacker News · 17h ago
Nissan's third generation e-POWER powertrain
mroche · Hacker News · 1d ago
MicroLLM Lab – Try 7 tiny LLM's in the browser
logicallee · Hacker News · 14h ago