Unsupervised Opinion Summarization as Copycat-Review Generation
Arthur Brazinskas, Mirella Lapata, Ivan Titov
Abstract
Opinion summarization is the task of automatically creating summaries that reflect subjective information expressed in multiple documents, such as product reviews. While the majority of previous work has focused on the extractive setting, i.e., selecting fragments from input reviews to produce a summary, we let the model generate novel sentences and hence produce abstractive summaries. Recent progress in summarization has seen the development of supervised models which rely on large quantities of document-summary pairs. Since such training data is expensive to acquire, we instead consider the unsupervised setting, in other words, we do not use any summaries in training. We define a generative model for a review collection which capitalizes on the intuition that when generating a new review given a set of other reviews of a product, we should be able to control the "amount of novelty" going into the new review or, equivalently, vary the extent to which it deviates from the input. At test time, when generating summaries, we force the novelty to be minimal, and produce a text reflecting consensus opinions. We capture this intuition by defining a hierarchical variational autoencoder model. Both individual reviews and the products they correspond to are associated with stochastic latent codes, and the review generator ("decoder") has direct access to the text of input reviews through the pointergenerator mechanism. Experiments on Amazon and Yelp datasets, show that setting at test time the review's latent code to its mean, allows the model to produce fluent and coherent summaries reflecting common opinions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ade36431-05b0-4cb5-83f5-b76ae00237f3Cited by top-tier papers25
- Factual and Informative Review Generation for Explainable RecommendationZhouhang Xie, Sameer Singh, Julian J. McAuley, Bodhisattwa Prasad MajumderAAAI 2023 · 36 citations
- Unsupervised Abstractive Dialogue Summarization for Tete-a-TetesXinyuan Zhang, Ruiyi Zhang, Manzil Zaheer, Amr AhmedAAAI 2021 · 27 citations
- Learning Opinion Summarizers by Selecting Informative ReviewsArthur Brazinskas, Mirella Lapata, Ivan TitovEMNLP 2021 · 26 citations
- Unsupervised Large Language Model Alignment for Information Retrieval via Contrastive FeedbackQian Dong, Yiding Liu, Qingyao Ai, Zhijing Wu et al.SIGIR 2024 · 9 citations
- Attributable and Scalable Opinion SummarizationTom Hosking, Hao Tang, Mirella LapataACL 2023 · 5 citations
Builds on1
Related papers
- Few-Shot Learning for Opinion SummarizationArthur Brazinskas, Mirella Lapata, Ivan TitovEMNLP 2020 · 4 citations
- Unsupervised Opinion Summarization with Noising and DenoisingReinald Kim Amplayo, Mirella LapataACL 2020 · 8 citations
- Unsupervised Extractive Opinion Summarization Using Sparse CodingSomnath Basu Roy Chowdhury, Chao Zhao, Snigdha ChaturvediACL 2022
- Unsupervised Opinion Summarisation in the Wasserstein SpaceJiayu Song, Iman Munire Bilal, Adam Tsakalidis, Rob Procter et al.EMNLP 2022 · 3 citations
- Unsupervised Opinion Summarization with Content PlanningReinald Kim Amplayo, Stefanos Angelidis, Mirella LapataAAAI 2021 · 51 citations
