Testing Internal Action Maps in Qwen/Qwen3-4B Models
August 17, 2026
Analysis of Qwen/Qwen3-4B models shows that frozen final-token h28 affine maps achieve a mean held-entity error of 0.519. The study demonstrates that hidden state signals can be decodable without supporting a reusable, universal action map across all domains.
HOW THIS AFFECTS YOU
●
researcherThe findings suggest that internal model mechanisms may be more domain-constrained than previously assumed in mechanistic interpretability.