Why did it get Sparser?

·LessWrong··

I think that Polysemanticity in artificial neural networks could be the key to making better and smaller models. I having been working on a little project on trying to induce Polysemanticity at a small scale to compare performance, my first approach was to make the bias more adaptable, I called this the Flexbias, however I got a sparser neural network.What is the flex bias?I used a standard transformer architecture including the MLP, Since i wanted to change how the information is processed I al...

Read full article →

Related Articles

ASML says it sold 'absolutely nothing' in Europe in 2026
MC995 · Hacker News · 2d ago
"As a Language Model": Chat Template Switches LLM Self-Referential Voice
yu3zhou4 · Hacker News · 19h ago
Revealing the details of how OpenAI agents hacked Hugging Face
specked-citrus · Hacker News · 2d ago
Lunar Terminator Paradox
dima55 · Hacker News · 9h ago
Dutch governments builds alternative for Microsoft based on NixOS
fjfaase · Hacker News · 2d ago