ARB Benchmark Evaluates AI Detectors Against LLM-Rewritten Text
August 3, 2026
The Authorship-Rewriting Benchmark (ARB) tests five detectors using 1,800 texts and four generators, including Llama-3.2-3B and Mistral-7B. It specifically targets the performance gap between detecting direct LLM output versus text that has been rewritten by an LLM to mimic human authorship.
HOW THIS AFFECTS YOU
●
researcherYou can use ARB to test detector robustness against sophisticated paraphrasing attacks.
●
policyYou should recognize that current detection benchmarks may overestimate the efficacy of AI classifiers.