Measuring Thinking Inertia and Instruction Non-Compliance in LLMs
October 9, 2026
Existing proxies for no-thinking modes, such as disabled reasoning traces, fail to capture actual model behavior. This study introduces three new metrics—Empty-Thinking Rate, Question-Pre-answer Relevance, and Explicit Inference Rate—to distinguish between true answer-only output and hidden or filler reasoning.
HOW THIS AFFECTS YOU
●
researcherYou can use these metrics to more accurately evaluate if models actually obey constraints to suppress reasoning.