Token-Level Detection of LLM Content in Coauthored Text
July 24, 2026
A new detection method uses an adaptive Lepski-type rule to smooth adjacent token scores, enabling localization of LLM-generated segments within human-AI collaborative documents. The approach identifies specific generated tokens without requiring token-level labeled training data.
HOW THIS AFFECTS YOU
●
researcherYou can evaluate the granularity of authorship in mixed-text datasets using this smoothing method.
●
policyThis helps in enforcing authenticity standards by pinpointing exactly which parts of a document are synthetic.