Can mid-training survive RL
About CaMLCaML is an alignment nonprofit working to make AI systems more compassionate toward all sentient beings with a focus on alignment midtraining (e.g. Teaching Claude Why). We build evaluation benchmarks on UK AISI's Inspect framework (TAC, MCP), run the public leaderboard at compassionbench.com, generate synthetic documents, and research which training interventions most effectively instill compassionate values. Our early results and those of others (Geodesic) show large gains from midtr...
Read full article →