arXiv · 2609.34030
Uncovering shortcut learning in audio classifiers by discovering recurring concepts in temporal explanations
Abstract
Correlations between events in machine learning datasets may result in shortcut learning, where models learn to predict the target event based on the presence of a correlated event. When these correlations are spurious -- arising from data collection artifacts -- models are likely to perform poorly in practice. We propose a pipeline to uncover shortcut learning in audio classifiers by discovering recurring concepts in their temporal explanations. Specifically, we isolate audio segments that explain classifier decisions, caption them with an ensemble of Large Audio-Language Models, and use a Large Language Model to extract recurring concepts. The resulting concepts can be audited by humans to uncover potential shortcut learning. We evaluate our framework using datasets curated from AudioSet Strong, controlling for the presence or absence of spurious correlations. Results show that this approach reliably uncovers learned shortcuts, such as the model relying on the presence of "laughter" to predict "applause".
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Cecilia Bolaños, Luciana Ferrer, Magdalena Fuentes. 2026-09-27. Uncovering shortcut learning in audio classifiers by discovering recurring concepts in temporal explanations. https://arxiv.org/abs/2609.34030
Cite the original work for its findings. Save a collection to share your selection of sources.