[arXiv]score: 0.24
Interpreting Language Model Hidden States at Scale
August 12, 2026
OmniLens scales trained lens methods to any model-width activation, including residual streams, attention, and MLP layers. Using low-rank translators, parameter growth becomes linear relative to model width, reducing trainable parameters by up to 98.4%. A Top-k Subset-KL training approach further decreases peak memory requirements by up to 70%.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy