What if Parameter Updates were Text?

·LessWrong··

Advice String Distillation This post will advocate for a fine-tuning methodology that I think is currently extremely under-rated for alignment and interpretability. It uses Context Distillation, but I don't think it has a proper name of its own yet, so I will refer to it as "Advice String Distillation" here. The goal is to accomplish the kind of fine-tuning that is done during RLVR, where models are trained to reliably carry out long chains of reasoning in order to accomplish tasks, without actu...

Read full article →

Related Articles

US sanctions force The Netherlands off Microsoft and toward alternative NixOS
mywacaday · Hacker News · 14h ago
How Delhi cut electricity loss from 50 to 5 percent
rbanffy · Hacker News · 13h ago
500k facial scans at UK stations yield no arrests, 1 false positive
ilamont · Hacker News · 14h ago
Livenerf: Has Opus 5.5 been nerfed yet?
bryan0 · Hacker News · 3h ago
A Privacy Analysis of Web and Mobile Conversational AI Agents [pdf]
damaru2 · Hacker News · 16h ago