Learning Opinion Summarizers by Selecting Informative Reviews
Arthur Brazinskas, Mirella Lapata, Ivan Titov
摘要
Opinion summarization has been traditionally approached with unsupervised, weaklysupervised and few-shot learning techniques. In this work, we collect a large dataset of summaries paired with user reviews for over 31,000 products, enabling supervised training. However, the number of reviews per product is large (320 on average), making summarization -and especially training a summarizerimpractical. Moreover, the content of many reviews is not reflected in the human-written summaries, and, thus, the summarizer trained on random review subsets hallucinates. In order to deal with both of these challenges, we formulate the task as jointly learning to select informative subsets of reviews and summarizing the opinions expressed in these subsets. The choice of the review subset is treated as a latent variable, predicted by a small and simple selector. The subset is then fed into a more powerful summarizer. For joint training, we use amortized variational inference and policy gradient methods. Our experiments demonstrate the importance of selecting informative reviews resulting in improved quality of summaries and reduced hallucinations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Factual and Informative Review Generation for Explainable RecommendationZhouhang Xie, Sameer Singh, Julian J. McAuley, Bodhisattwa Prasad MajumderAAAI 2023 · 被引用 36 次
- Shilling Black-box Review-based Recommender Systems through Fake Review GenerationHung-Yun Chiang, Yi-Syuan Chen, Yun-Zhu Song, Hong-Han Shuai 等KDD 2023 · 被引用 15 次
- Attributable and Scalable Opinion SummarizationTom Hosking, Hao Tang, Mirella LapataACL 2023 · 被引用 5 次
- Less Is More? Examining Fairness in Pruned Large Language Models for Summarising OpinionsNannan Huang, Haytham M. Fayek, Xiuzhen ZhangEMNLP 2025 · 被引用 2 次
- How to Compare Things Properly? A Study of Argument Relevance in Comparative Question AnsweringIrina Nikishina, Saba Anwar, Nikolay Dolgov, Maria Manina 等ACL 2025
它引用的顶会 Paper9
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Dice Loss for Data-imbalanced NLP TasksXiaoya Li, Xiaofei Sun, Yuxian Meng, Junjun Liang 等ACL 2020 · 被引用 575 次
- Asking and Answering Questions to Evaluate the Factual Consistency of SummariesAlex Wang, Kyunghyun Cho, Mike LewisACL 2020 · 被引用 317 次
- Pre-training via ParaphrasingMike Lewis, Marjan Ghazvininejad, Gargi Ghosh, Armen Aghajanyan 等NeurIPS 2020 · 被引用 165 次
- On Faithfulness and Factuality in Abstractive SummarizationJoshua Maynez, Shashi Narayan, Bernd Bohnet, Ryan T. McDonaldACL 2020 · 被引用 54 次
相关 Paper
- Unsupervised Opinion Summarization as Copycat-Review GenerationArthur Brazinskas, Mirella Lapata, Ivan TitovACL 2020 · 被引用 14 次
- Few-Shot Learning for Opinion SummarizationArthur Brazinskas, Mirella Lapata, Ivan TitovEMNLP 2020 · 被引用 4 次
- Unsupervised Opinion Summarization with Noising and DenoisingReinald Kim Amplayo, Mirella LapataACL 2020 · 被引用 8 次
- Unsupervised Opinion Summarization with Content PlanningReinald Kim Amplayo, Stefanos Angelidis, Mirella LapataAAAI 2021 · 被引用 51 次
- Unsupervised Extractive Opinion Summarization Using Sparse CodingSomnath Basu Roy Chowdhury, Chao Zhao, Snigdha ChaturvediACL 2022
