[arXiv]score: 0.18
Shared circuits predict whether LLMs generalize across formats in arithmetic reasoning
September 7, 2026
Attribution patching identifies distinct circuits for numeric versus verbal arithmetic reasoning across English, Spanish, and Italian. Overlap between these task-specific circuits predicts a model's ability to generalize across input formats. This suggests that internal circuit commonality serves as a structural proxy for cross-format robustness.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy