Language models with as few as 410 million parameters match human performance in tracking entities through naturalistic narratives. Performance scales beyond human capabilities in contemporary large-scale models, demonstrating that complex discourse comprehension does not require multi-billion parameter architectures.
HOW THIS AFFECTS YOU
●
builderYou can achieve high-quality narrative comprehension using much smaller, more efficient models.
●
researcherThis suggests entity tracking is an emergent property that scales predictably even in small models.