SFT-as-Context Reduces Forgetting in Specialized LLM Fine-Tuning
October 9, 2026
SFT-as-context is a training-free method that prevents catastrophic forgetting by using a fine-tuned model's response as context for the original parent model. In testing, it maintained general capabilities within 2.2 percentage points of the parent model while remaining competitive on specialized benchmarks like AIME 2024.
HOW THIS AFFECTS YOU
●
builderYou can deploy specialized models without losing the reasoning or general capabilities of your base model.
●
researcherThis offers a new training-free paradigm for studying the trade-offs between specialization and generalization.