Shared Fisher-Rao Geometry Across LLM Architectures
September 11, 2026
Analysis shows that the Fisher-Rao geometry of next-token probabilities is shared across transformer, state-space, and recurrent models. This output geometry correlates with semantic-category transfer and predictive accuracy, independent of specific activation coordinates.
HOW THIS AFFECTS YOU
●
researcherYou can leverage shared output geometry to predict semantic-category transfer across different model architectures.