Impact of Lie Typology and Sparsity on LLM Deception Detection
July 24, 2026
A systematic study reveals that deception detection probes are highly sensitive to lie typology, representation depth, and sparsity. Results show that training on limited lie types fails to generalize and that optimal probe depth is heavily dependent on the specific dataset used.
HOW THIS AFFECTS YOU
●
researcherYou should account for lie typology and representation depth when designing probes for deceptive output detection.
●
policyThis highlights the difficulty of creating reliable automated safeguards against model deception.