Large-Scale Correlation Analysis of Automated Metrics for Topic Models
Jia Peng Lim, Hady W. Lauw
Abstract
Automated coherence metrics constitute an important and popular way to evaluate topic models. Previous works present a mixed picture of their presumed correlation with human judgement. In this paper, we conduct a large-scale correlation analysis of coherence metrics. We propose a novel sampling approach to mine topics for the purpose of metric evaluation, and conduct the analysis via three large corpora showing that certain automated coherence metrics are correlated. Moreover, we extend the analysis to measure topical differences between corpora. Lastly, we examine the reliability of human judgement by conducting an extensive user study, which is designed as an amalgamation of different proxy tasks to derive a finer insight into the human decision-making processes. Our findings reveal some correlation between automated coherence metrics and human judgement, especially for generic corpora. C γ=1 V,̸ e C γ=2 V,̸ e C NPMI,̸ e C NPMI C P,o C UMass,o C γ=1 V,̸ e -0.87 0.95 0.74 0.81 0.33 C γ=2 V,̸ e
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0295c9a0-805f-4404-89d7-b8d45aa9074bCited by top-tier papers3
- PromptMTopic: Unsupervised Multimodal Topic Modeling of Memes using Large Language ModelsNirmalendu Prakash, Han Wang, Nguyen-Khoi Hoang, Ming Shan Hee et al.ACM MM 2023 · 24 citations
- Enhancing Topic Interpretability for Neural Topic Modeling Through Topic-Wise Contrastive LearningXin Gao, Yang Lin, Ruiqing Li, Yasha Wang et al.ICDE 2024 · 3 citations
- Disentangling Transformer Language Models as Superposed Topic ModelsJia Peng Lim, Hady W. LauwEMNLP 2023 · 2 citations
Builds on12
- Is Automated Topic Model Evaluation Broken? The Incoherence of CoherenceAlexander Miserlis Hoyle, Pranav Goel, Andrew Hian-Cheong, Denis Peskov et al.NeurIPS 2021 · 220 citations
- With Little Power Comes Great ResponsibilityDallas Card, Peter Henderson, Urvashi Khandelwal, Robin Jia et al.EMNLP 2020 · 76 citations
- Graph Attention Topic Modeling NetworkLiang Yang, Fan Wu, Junhua Gu, Chuan Wang et al.WWW 2020 · 57 citations
- Hierarchical Topic Mining via Joint Spherical Tree and Text EmbeddingYu Meng, Yunyi Zhang, Jiaxin Huang, Yu Zhang et al.KDD 2020 · 56 citations
- Topic Modeling Revisited: A Document Graph-based Neural Network PerspectiveDazhong Shen, Chuan Qin, Chao Wang, Zheng Dong et al.NeurIPS 2021 · 50 citations
Related papers
- Evaluation of Thematic Coherence in MicroblogsIman Munire Bilal, Bo Wang, Maria Liakata, Rob Procter et al.ACL 2021
- Evaluating Dynamic Topic ModelsCharu James, Mayank Nagda, Nooshin Haji Ghassemi, Marius Kloft et al.ACL 2024 · 1 citation
- ProxAnn: Use-Oriented Evaluations of Topic Models and Document ClusteringAlexander Miserlis Hoyle, Lorena Calvo-Bartolomé, Jordan Lee Boyd-Graber, Philip ResnikACL 2025
- Beyond correlation: The impact of human uncertainty in measuring the effectiveness of automatic evaluation and LLM-as-a-judgeAparna Elangovan, Lei Xu, Jongwoo Ko, Mahsa Elyasi et al.ICLR 2025
- User Ex Machina : Simulation as a Design Probe in Human-in-the-Loop Text AnalyticsAnamaria Crisan, Michael CorrellCHI 2021 · 11 citations
