Posshbench Evaluates Neural Model Generalization from Limited Stimuli
August 17, 2026
The posshbench framework evaluates Transformer, LSTM, and n-gram models on four canonical linguistic phenomena. Findings show models can achieve above-chance generalization from 10M words of input, though they remain less efficient learners than humans as scale increases.
HOW THIS AFFECTS YOU
●
researcherThis provides a unified benchmark to test if architectural inductive biases improve learning efficiency.