LLM agent credit assignment signals fail to outperform chance
August 21, 2026
Auditing step-level credit signals—including LLM-judge scores and logprob ratios—against causal ground truth in ALFWorld shows these methods fail to identify causally significant steps better than chance. Causal contribution is found to be sparse, affecting only 30.5% of decision points.
HOW THIS AFFECTS YOU
●
researcherYou should prioritize causal intervention testing over correctness-based evaluation for agent training.