PaperScope
LIVE · 2026-09-10 05:40 UTC

Beyond Top Words: MonoTM for Topic Modeling with Interpretable Monosemantic Features

Una Joh, Bei Yu

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.09575 v1
Category
Submitted
2026-09-09

Abstract

Topic models summarize large text corpora, but top-ranked words often provide only a limited representation of topic semantics. Sparse autoencoders (SAEs) offer a way to move beyond word-level descriptors by extracting interpretable features from dense representations, yet how feature interpretability relates to topic-inference quality remains unclear. We introduce \textbf{MonoTM}, an interpretable topic modeling framework that decouples these roles. Across three benchmark corpora, we show that document--topic mixture estimation and semantic interpretation favor different SAE configurations and feature subsets. MonoTM estimates mixtures from the full SAE bag-of-features representation and, with them fixed, learns topic descriptors over a separate vocabulary of corpus-grounded semantic features. This design preserves global topic structure while representing topics with semantic units more meaningful than individual words, making them more useful for downstream corpus analysis.

Comment: Accepted to appear in the Proceedings of AACL-IJCNLP 2026

arXiv abs page · PDF · same-day batch