PsyAgentBench: Separating Pattern Recognition from Psychological Bias in LLMs
September 22, 2026
PsyAgentBench uses a factorial design of named vs. blind prompts and canonical vs. counterfactual tasks to distinguish between LLM pattern matching and true psychological bias. Testing on open-weight models shows effects like Asch conformity are often triggered by paradigm-label gating rather than inherent model susceptibility.
HOW THIS AFFECTS YOU
●
researcherYou can use this benchmark to determine if model behaviors are true biases or just training data contamination.
●
policyThis provides evidence for how model alignment might be masking underlying behavioral vulnerabilities.