Sparse Autoencoders Reveal Neural Differences Between CoT and Direct Generation
August 11, 2026
Applying Top-K Sparse Autoencoders to DeepSeek-R1-Distill-Qwen-7B shows that Thinking mode relies on sparse, high-intensity feature activations for verbal deduction, whereas NoThinking mode uses diffuse patterns for symbolic manipulation. Suppressing the three most active features reveals how reasoning mechanisms diverge across task complexities.
HOW THIS AFFECTS YOU
●
researcherYou can use sparse autoencoders to interpret the specific mechanistic differences in Chain-of-Thought reasoning.