LLMs Diverge from Human Standards in Creativity Evaluation
July 27, 2026
Analysis of six LLMs reveals that while models align with humans on novelty, they diverge significantly on contextual dimensions like social and market information. Each model uses distinct, narrow evaluation standards that affect the consistency of creativity judgments.
HOW THIS AFFECTS YOU
●
researcherYou should be cautious when using LLMs as proxies for human creative assessment due to narrow standard adoption.
●
designerThis affects how you can use AI to critique or guide creative workflows.