NB-LoRA Preserves Reasoning During Post-RL Fine-Tuning
September 23, 2026
NB-LoRA identifies low-dimensional subspaces where reasoning activations concentrate to allow parameter-efficient adaptation without overwriting RL-elicited capabilities. It estimates approximate null spaces from minimal examples to mitigate catastrophic forgetting during supervised fine-tuning.
HOW THIS AFFECTS YOU
●
builderYou can adapt reasoning models to new domains more safely using this method.
●
researcherYou can explore how null-space capacity facilitates domain adaptation without losing reasoning skills.