Can mid-training survive RL

·LessWrong··

About CaMLCaML is an alignment nonprofit working to make AI systems more compassionate toward all sentient beings with a focus on alignment midtraining (e.g. Teaching Claude Why). We build evaluation benchmarks on UK AISI's Inspect framework (TAC, MCP), run the public leaderboard at compassionbench.com, generate synthetic documents, and research which training interventions most effectively instill compassionate values. Our early results and those of others (Geodesic) show large gains from midtr...

Read full article →

Related Articles

The ChatGPT/Codex app bundles a full copy of LibreOffice
timpera · Hacker News · 14h ago
The Emergent Symbolic Structure of Artificial Neural Networks
schmuhblaster · Hacker News · 5h ago
I trained a small transformer in 1.5hrs and it beats many LLMs
porridgeraisin · Hacker News · 1d ago
Show HN: Running 104GB Qwen3.8-Flash-Next on 48GB Mac with at ~12 tok/s
carloslfu · Hacker News · 17h ago
Refurbishing a Tektronix TDS7104 Oscilloscope
jwise0 · Hacker News · 14h ago