LLM Agent Groups Overstate Consensus in Reasoning Tasks
September 18, 2026
A study replaying human Wason reasoning tasks with LLM agents shows that AI groups exhibit significantly higher consensus rates than humans. Even when participation is matched, agent groups show consensus gaps of up to 44.4 percentage points, suggesting a tendency toward groupthink in multi-agent reasoning.
HOW THIS AFFECTS YOU
●
builderAvoid relying on agent consensus as a proxy for truth or collective intelligence in your applications.
●
researcherYou should account for artificial consensus bias when evaluating multi-agent debate or reasoning frameworks.