Is there even a ground-truth for LLMs’ internal representations?

·LessWrong··

[This is an introductory blog for the paper Laguerre Geometry for Interpreting Large Language Models and the GitHub repository Geometric Lens.]LLM Lens: What does an internal vector mean?Anthropic's recent paper on the "J-Lens" (Jacobian Lens) has revived interest in reading the "thoughts" inside Large Language Models. The idea of placing a “lens” at an LLM's hidden layers isn't new. It dates back to the Logit Lens, and has since evolved into a family of variants, including Tuned Lens and Patchs...

Read full article →

Related Articles

Hacker wipes Romania's land registry database
speckx · Hacker News · 9h ago
Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
cl42 · Hacker News · 7h ago
How we measured AI writing across arXiv, and where the measurement breaks
dopamine_daddy · Hacker News · 6h ago
Claude Code uses Bun written in Rust now
tosh · Hacker News · 1d ago
Xiaomi-Robotics-1
ilreb · Hacker News · 18h ago