One message is all it takes: a failure of critical thinking in LLMs

·LessWrong··

summary: For a while I've suspected that modern LLMs are getting better at solving posed problems, while progress in critical thinking stagnates, or even regresses, losing the ability to judge the meaning of a result. The July counterexample to the Jacobian conjecture is a rare way to test that: a huge prior overturned by something a model can verify by itself in one reply. I let the model verify the counterexample itself, then gaslight it with a single message of about ten words. The model drop...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 13h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 1d ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 8h ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 6h ago
Nvidia announces native GPU programming in Rust
nonmaskable · Hacker News · 15h ago