Controllable-CoT leads to covert reasoning capabilities

·LessWrong··

SummaryI measure GPT-6 Astra’s performance on multi-hop tasks when prompted with a secondary CoT-control instruction: to reason using only dots, or to reason steganographically. Astra demonstrates covert reasoning capabilities with task performance beating that when using no reasoning or filler tokens for reasoning.This work agrees with findings from Astra is much better at reasoning with filler tokens than previous models but has the model generate its own reasoning and provide it as part of th...

Read full article →

Related Articles

NASA’s Mars Sample Return mission is dead
Muhammad523 · Hacker News · 20h ago
AMD's random number generator can't generate a 0?
BruceEel · Hacker News · 7h ago
What happened to the Snowden archive
EXHades · Hacker News · 1d ago
Samsung is expected to more than double output of its HBM4 and HBM4E DRAM
giuliomagnifico · Hacker News · 1d ago
MiMo-v2.6-Pro: Intelligence, Performance and Price Analysis
theanonymousone · Hacker News · 12h ago