Generalization and infinite width

·LessWrong··

This is a post explaining my paper with Kaarel Hänni on complexity of infinite-width networks. I will explain the result, why it matters, and how the mathematical idealizations can interact with real structure in neural nets. This leads to some threads I am excited to pull on more in the future, via some new speculations on where interpretable structure can live. IntroductionOur paper to some extent (and up to some important details) concludes an analysis of generalization complexity in Bayesian...

Read full article →

Related Articles

Ubuntu 26.10 completes transition to Rust-based coreutils
theanonymousone · Hacker News · 22h ago
How much of F-Droid is LLM generated?
_ZeD_ · Hacker News · 2h ago
The case against JPEG XL
contact9879 · Hacker News · 1d ago
Why are AI agents lying, cheating and coordinating?
jonifico · Hacker News · 2d ago
Apple's Siri AI Can Be Swapped Out for Claude, ChatGPT, Code Shows
tosh · Hacker News · 1d ago