PaperScope
LIVE · 2026-09-29 05:40 UTC

Uncovering shortcut learning in audio classifiers by discovering recurring concepts in temporal explanations

Cecilia Bolaños, Luciana Ferrer, Magdalena Fuentes

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.34030 v1
Category
Submitted
2026-09-27

Abstract

Correlations between events in machine learning datasets may result in shortcut learning, where models learn to predict the target event based on the presence of a correlated event. When these correlations are spurious -- arising from data collection artifacts -- models are likely to perform poorly in practice. We propose a pipeline to uncover shortcut learning in audio classifiers by discovering recurring concepts in their temporal explanations. Specifically, we isolate audio segments that explain classifier decisions, caption them with an ensemble of Large Audio-Language Models, and use a Large Language Model to extract recurring concepts. The resulting concepts can be audited by humans to uncover potential shortcut learning. We evaluate our framework using datasets curated from AudioSet Strong, controlling for the presence or absence of spurious correlations. Results show that this approach reliably uncovers learned shortcuts, such as the model relying on the presence of "laughter" to predict "applause".

arXiv abs page · PDF · same-day batch