Evaluation of Thematic Coherence in Microblogs
Iman Munire Bilal, Bo Wang, Maria Liakata, Rob Procter, Adam Tsakalidis
摘要
Collecting together microblogs representing opinions about the same topics within the same timeframe is useful to a number of different tasks and practitioners. A major question is how to evaluate the quality of such thematic clusters. Here we create a corpus of microblog clusters from three different domains and time windows and define the task of evaluating thematic coherence. We provide annotation guidelines and human annotations of thematic coherence by journalist experts. We subsequently investigate the efficacy of different automated evaluation metrics for the task. We consider a range of metrics including surface level metrics, ones for topic model coherence and text generation metrics (TGMs). While surface level metrics perform well, outperforming topic coherence metrics, they are not as consistent as TGMs. TGMs are more reliable than all other metrics considered for capturing thematic coherence in microblog clusters due to being less sensitive to the effect of time windows.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- Short Text Topic Modeling with Topic Distribution Quantization and Negative Sampling DecoderXiaobao Wu, Chunping Li, Yan Zhu, Yishu MiaoEMNLP 2020 · 被引用 61 次
- BLEURT: Learning Robust Metrics for Text GenerationThibault Sellam, Dipanjan Das, Ankur P. ParikhACL 2020 · 被引用 40 次
相关 Paper
- Large-Scale Correlation Analysis of Automated Metrics for Topic ModelsJia Peng Lim, Hady W. LauwACL 2023 · 被引用 13 次
- Is Automated Topic Model Evaluation Broken? The Incoherence of CoherenceAlexander Miserlis Hoyle, Pranav Goel, Andrew Hian-Cheong, Denis Peskov 等NeurIPS 2021 · 被引用 220 次
- Evaluating Dynamic Topic ModelsCharu James, Mayank Nagda, Nooshin Haji Ghassemi, Marius Kloft 等ACL 2024 · 被引用 1 次
- ProxAnn: Use-Oriented Evaluations of Topic Models and Document ClusteringAlexander Miserlis Hoyle, Lorena Calvo-Bartolomé, Jordan Lee Boyd-Graber, Philip ResnikACL 2025
- Large-Scale Evaluation of Topic Models and Dimensionality Reduction Methods for 2D Text SpatializationDaniel Atzberger, Tim Cech, Matthias Trapp, Rico Richter 等IEEE VIS 2023 · 被引用 11 次
