You can graft SDF changes from base models onto post-trained models
Webpage: praxis-research.org/grafting | Paper link: https://arxiv.org/abs/2610.00767Work done during Peter and Dani's MATS Fellowship (10.0) under the mentorship of Shi, with Jinghua Ou as Research Manager. TLDR;We share a simple modification to synthetic document fine-tuning (SDF) that works surprisingly well: do fine-tuning on a pre-training checkpoint and add the resulting weight difference onto the post-trained version that you want to deploy. We call this grafting.We empirically validate gr...
Read full article →