Discovering Hidden Causal Interfaces in Language Models
July 31, 2026
The forked futures method identifies reusable causal interfaces in LLM hidden states by comparing response distributions induced by future operations. The 'Shared' interface architecture achieved gains of up to 0.294 nats on Llama-3-8B.
HOW THIS AFFECTS YOU
●
researcherYou can identify and exploit reusable internal model interfaces without requiring researcher-specified latent labels.