Exploration: fine-tuning with parameter decomposition

·LessWrong··

TL;DR: We can destroy a 67M-parameter language model's ability to predict German text by fine-tuning a single number: the scalar prefactor on one German-related rank-1 parameter subcomponent. This is an early exploration into using parameter decomposition for a more targeted and interpretable form of model fine-tuning. At small German-token budgets, fine-tuning the scalar prefactor of a single German-related parameter subcomponent beats rank-1 and rank-4 LoRA [1] fine-tunes on the trade-off betw...

Read full article →

Related Articles

Italian parliament votes for return to nuclear energy
geox · Hacker News · 4h ago
Claude Opus 5.5
km144 · Hacker News · 1d ago
GPT-6 Astra has gained the ability to drive a car
plurby · Hacker News · 6h ago
UK military jamming other nations' satellites to defend itself, BBC told
thm · Hacker News · 4h ago
Claude Code reads AGENTS.md only when telemetry is on [fixed]
pszypowicz · Hacker News · 9h ago