Latent reasoning architectures would likely undermine CoT, our strongest oversight tool

·Redwood Research··

Edited 2026-09-29; see note at the start of the conclusion.Summary: Currently, “chain of thought” (CoT) is our most valuable tool for understanding the reasoning and cognition of AI systems. However, some architectures would enable AI models to reason much more extensively in latent states rather than in text CoT. We think that a shift towards latent reasoning architectures would likely undermine the usefulness of CoT and make oversight much harder.IntroductionSwarms of more than a thousand AI a...

Read full article →

Related Articles

MIT's New Method Flags AI Models Trained on CASM Without Generating It
sdoering · Hacker News · 2mo ago
Harm Laundering in GPT Models: Gender Discrimination Transformed Rather Than
sbulaev · Hacker News · 10d ago
Continual learning might make your blocking monitors nearly useless
Alex Mallen · Alignment Forum · 4d ago
Can parts of the HuggingFace incident be simulated?
Benedikt Droste · LessWrong · 12d ago
The Hobbesian Bootstrap Paradox in Frontier AI
Claudio Di Meglio · EA Forum · 15d ago