Decoupling of Internal and Verbalized Probabilities in Large Language Models
October 2, 2026
Research shows that an LLM's internal sampling distribution and its verbalized confidence are not inherently aligned. The degree of alignment depends on whether training data frequencies match explicit probabilistic assertions, limiting the reliability of using verbalized uncertainty as a proxy for model distribution.
HOW THIS AFFECTS YOU
●
researcherYou should be cautious when using model self-reports as a proxy for actual calibration or sampling confidence.