VLMs Fail to Express Uncertainty Despite Internal Knowledge
August 14, 2026
The TRAPSBench benchmark reveals that while vision-language models can internally identify when visual evidence is insufficient, they fail to express this uncertainty. Linear probes show models decode answerability with up to 0.91 AUROC, yet they rarely abstain from answering.
HOW THIS AFFECTS YOU
●
researcherYou should investigate the gap between internal latent perception and external verbal expression in multimodal models.
●
designerThis highlights the need for UX patterns that handle model overconfidence in uncertain scenarios.