Evidence about risk should be transparent by Ajeya

·Nuno Sempere··

Note: This post was cross­posted from Planned Ob­so­les­cence by the Fo­rum team, with the au­thor’s per­mis­sion. The au­thor may not see or re­spond to com­ments on this post.Subti­tle: We can’t de­velop safety stan­dards if we have to rely on opaque judgmentAll views are my own and do not rep­re­sent my em­ployer.In the wake of the re­cent wave of mis­al­ign­ment in­ci­dents, both OpenAI and An­thropic have re­ported slow­ing down RL train­ing to im­prove safety. Th­ese in­ci­dents, com­bined...

Read full article →

Related Articles

Will my flight be delayed | will I miss my connection?
Genzy · Manifold Markets · 23h ago
Will one AI lab pull ahead before 2028?
SG · Manifold Markets · 1d ago
Gas prices in the US on September 30 2026?
Fastcar99 · Manifold Markets · 3d ago
Will Opus 5.5 get nerfed within a week of launch?
Hi · Manifold Markets · 3d ago
AI safety field *visual* impact analysis by hannarchy
hannarchy · Nuno Sempere · 4d ago