PaperScope
LIVE · 2026-09-03 05:40 UTC

When Less is More: Understanding When Token Filtering Helps and Fails in AI-generated Text Detection

Xiaoyang Han, Lvxiaowei Xu, Ming Cai

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2608.29903 v1
Category
Submitted
2026-08-30

Abstract

The rapid advancement of large language models (LLMs) has made AI-generated text detection increasingly critical. Existing zero-shot detectors assume that more token-level evidence leads to more reliable detection. However, our empirical study challenges this consensus: fewer tokens sometimes work better, retaining only 40% can yield optimal performance, yet this benefit is not universal. Using the Entropy Gap Score (EGS), we introduce top-$k$ cumulative probability filtering as a diagnostic probe. Across three representative settings, filtering exhibits strikingly different behaviors. We analyze EGS via typical set theory and quantify its dynamics through entropy calibration and distribution analysis. We find that filtering helps for weak source LMs, where low-entropy tokens are harmful, but fails for strong source LMs, where they are not notably harmful. Our work provides the first systematic analysis showing that some tokens are not merely uninformative but systematically harmful due to entropy miscalibration, revealing a two-sided trade-off in token-level detection.

Comment: Accepted to EMNLP 2026 (Main Conference)

arXiv abs page · PDF · same-day batch