LLM detectors suffer 76.3% false-positive rate on formal human text
October 5, 2026
Analysis of a RoBERTa-based detector reveals that increasing verb diversity in Mistral-7B-Instruct outputs actually makes them easier to detect. The study finds detection scores track statistical complexity rather than machine-likeness, resulting in high false-positive rates for formal human writing and low robustness to paraphrasing.
HOW THIS AFFECTS YOU
●
researcherYou should consider the high error rates and lack of structural robustness in current supervised AI-text detectors.
●
policyThis highlights the unreliability of automated detectors for governing AI-generated content.