PaperScope
LIVE · 2026-09-03 05:40 UTC

Disentangling Statistical Preemption from Entrenchment in Language Models' Avoidance of Overgeneralization

Yixuan Wang, Freda Shi, Kanishka Misra

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.01794 v1
Category
Submitted
2026-09-01

Abstract

How do learners avoid overgeneralizations such as Tom laughed me without explicit negative evidence? Constructionists have posited two proposals that describe indirect negative evidence against overgeneralizations: preemption (which privileges exposure to near-synonymous construction---e.g., she made him laugh) vs. entrenchment (all exposures to a verb's grammatical usages, including cases like He laughed). We disentangle these hypotheses by running controlled rearing experiments on LMs trained on child-caregiver conversations, where we systematically remove preemptive vs. non-preemptive evidence. We find that while LMs avoid overgeneralizations, they do not show preemption at a verb-specific level, instead showing weak but non-zero evidence of abstract preemption. Combined with results from analyzing the LMs' training dynamics, we find that LMs treat competing structures as indirect positive---as opposed to negative---evidence in the verb-specific condition. Insofar as preemption is the more plausible route to avoiding overgeneralizations in humans, our results point the need for there to be sensitivities to indirect negative evidence in neural network learners, and suggest new human experiments to test abstract preemption.

Comment: Accepted to EMNLP 2026 Main

arXiv abs page · PDF · same-day batch