Thought anchors don't transfer between models

·LessWrong··

TL; DR. Sentences deemed important in the CoT trace of one model are very ordinary when a different reasoning model reads the same trace. Their importance is model-specific and not a property of the text itself. In contrast, a hidden nudge translates across models: a CoT written by a model that silently followed a hint has a significant impact when teacher-forced into another model. However, it only works if the CoT trace already discusses the answer options. Otherwise, it has no effect within t...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 10h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 21h ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 5h ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 3h ago
Building a Linux GPU Driver for the M4 Mac Mini in One Month
ADevWithAnIdea · Hacker News · 1d ago