ReCal Improves Reasoning Recovery in Pruned Models
October 9, 2026
ReCal uses forward KL divergence between unpruned teachers and pruned probes to adjust calibration before structured pruning. This approach improves on-policy distillation recovery, yielding up to a 16.7 percentage point gain on AIME mathematical reasoning benchmarks.
HOW THIS AFFECTS YOU
●
builderYou can use this to deploy smaller, pruned reasoning models that retain more of their teacher's mathematical capabilities.
●
researcherThe method provides a way to guide pruning criteria toward preserving high-value teacher predictions.