Internal State Control is a General Property of LLMs

·LessWrong··

tl;dr:Lindsey 2025 found models can modulate their internal states: when instructed to “think about” a concept while writing an unrelated sentence, the representation of the concept is more present than when instructed to not think about it.Internal state controllability appears to be a general property of LLMs: the effect replicates in 14 open-weight models from 0.3B to 235B parameters (Qwen3, Gemma 3, Tulu 3) with no clear trend in the think vs. don't-think gap across scale.Since controllabili...

Read full article →

Related Articles

AI's top startups are barely publishing their research
YeGoblynQueenne · Hacker News · 1d ago
Document-borne AI worms can self-propagate through Copilot for Word
Canopy9560 · Hacker News · 1d ago
Why Don't People Use Formal Methods? (2019)
Thom2503 · Hacker News · 9h ago
Keychron announces first open-source firmware for gaming mice
JLO64 · Hacker News · 1d ago
Turning a dumb AC unit smart (without losing my security deposit)
austinallegro · Hacker News · 1d ago