Language Model Word Order Bias Driven by Data Density
August 18, 2026
Transformer models exhibit a preference for right-branching SVO word orders in natural languages, correlating with high-resource data availability rather than innate architectural bias. On artificial languages, models instead show a left-branching preference, proving that observed word order biases are data-driven rather than structural.
HOW THIS AFFECTS YOU
●
researcherYou should account for data distribution when evaluating architectural inductive biases in multilingual training.