SAGO Framework Measures LLM Generalization via Semantic Stability
October 2, 2026
The Stability-Aware Generalization Objective (SAGO) framework shifts evaluation from aggregate accuracy to individual example stability. It measures how much model outputs vary when the same input is expressed through different semantic variations, preventing performance optimization via narrow training.
HOW THIS AFFECTS YOU
●
researcherYou can use this to detect if models are truly generalizing or merely memorizing specific prompt formats.