Is there even a ground-truth for LLMs’ internal representations?

·LessWrong··

[This is an introductory blog for the paper Laguerre Geometry for Interpreting Large Language Models and the GitHub repository Geometric Lens.]LLM Lens: What does an internal vector mean?Anthropic's recent paper on the "J-Lens" (Jacobian Lens) has revived interest in reading the "thoughts" inside Large Language Models. The idea of placing a “lens” at an LLM's hidden layers isn't new. It dates back to the Logit Lens, and has since evolved into a family of variants, including Tuned Lens and Patchs...

Read full article →

Related Articles

OpenAI's GPT-6 Astra on ARC-AGI-3
vignesh_warar · Hacker News · 8h ago
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out
screm · Hacker News · 7h ago
GLP-1s are being linked to fewer serious infections, including TB
gumby · Hacker News · 5h ago
Pre-Release of Polars 2.0
komape · Hacker News · 21h ago
Artificial beaver dams saw juvenile coho salmon survival rates go from 8% to 60%
speckx · Hacker News · 12h ago