Mechanistic Analysis of Value Reading in Qwen2.5VL and InternVL
September 15, 2026
Counterfactual activation patching on Qwen2.5VL-7B and InternVL3.5-8B reveals that bar-top regions contribute more to value recovery than bar bodies. The study traces how models shift from using visual legend regions in early layers to prompt-series states in middle layers.
HOW THIS AFFECTS YOU
●
researcherThis clarifies how VLMs process spatial geometry versus textual legends.
●
designerUnderstanding these attention patterns can inform better chart visualization for multimodal models.