HouseholdBench Evaluates LLM Accuracy in Predicting U.S. Economic Behavior
October 7, 2026
HouseholdBench tests 13 proprietary and open-weight models across 6 U.S. surveys and 32 prediction tasks involving consumption, income, and labor. Results show LLMs can outperform no-change baselines in predicting how households adjust to macroeconomic policy changes.
HOW THIS AFFECTS YOU
●
researcherYou can use this benchmark to evaluate how well models simulate human socio-economic responses.
●
investorThis highlights the potential for LLMs to serve as quantitative proxies for consumer behavior in economic modeling.