Reflecting on Experiences for Response Generation
Chenchen Ye, Lizi Liao, Suyu Liu, Tat-Seng Chua
Abstract
Multimodal dialogue systems attract much attention recently, but they are far from skills like: 1) automatically generate context- specific responses instead of safe but general responses; 2) naturally coordinate between the different information modalities (e.g. text and image) in responses; 3) intuitively explain the reasons for generated responses and improve a specific response without re-training the whole model. To approach these goals, we propose a different angle for the task - Reflecting Experiences for Response Generation (RERG). This is supported by the fact that generating a response from scratch can be hard, but much easier if we can access other similar dialogue contexts and the corresponding responses. In particular, RERG first uses a multimodal contrastive learning enhanced retrieval model for soliciting similar dialogue instances. It then employs a cross copy based reuse model to explore the current dialogue context (vertical) and similar dialogue instances' responses (horizontal) for response generation simultaneously. Experimental results demonstrate that our model outperforms other state-of-the-art models on both automatic metrics and human evaluation. Moreover, RERG naturally provides supporting dialogue instances for better explainability. It also has a strong capability in adapting to unseen dialogue settings by simply adding related samples to the retrieval datastore without re-training the whole model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5c124e90-08b4-4bbc-bdfd-97da14f201f5Cited by top-tier papers5
- Broadening the View: Demonstration-augmented Prompt Learning for Conversational RecommendationHuy Dao, Yang Deng, Dung D. Le, Lizi LiaoSIGIR 2024 · 19 citations
- Reinforced Target-driven Conversational PromotionHuy Dao, Lizi Liao, Dung D. Le, Yuxiang NieEMNLP 2023 · 3 citations
- Retrieval Augmented Generation for Dynamic Graph ModelingYuxia Wu, Lizi Liao, Yuan FangSIGIR 2025 · 2 citations
- Self-chats from Large Language Models Make Small Emotional Support Chatbot BetterZhonghua Zheng, Lizi Liao, Yang Deng, Libo Qin et al.ACL 2024
- Thoughts to Target: Enhance Planning for Target-driven ConversationZhonghua Zheng, Lizi Liao, Yang Deng, Ee-Peng Lim et al.EMNLP 2024
Builds on14
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Unified Vision-Language Pre-Training for Image Captioning and VQALuowei Zhou, Hamid Palangi, Lei Zhang, Houdong Hu et al.AAAI 2020 · 1,047 citations
- Generalization through Memorization: Nearest Neighbor Language ModelsUrvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer et al.ICLR 2020 · 1,038 citations
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz et al.NeurIPS 2020 · 590 citations
Related papers
- Multimodal Dialogue Response GenerationQingfeng Sun, Yujing Wang, Can Xu, Kai Zheng et al.ACL 2022 · 58 citations
- ZRIGF: An Innovative Multimodal Framework for Zero-Resource Image-Grounded Dialogue GenerationBo Zhang, Jian Wang, Hui Ma, Bo Xu et al.ACM MM 2023 · 4 citations
- Conversational Composed Retrieval with Iterative Sequence RefinementHao Wei, Shuhui Wang, Zhe Xue, Shengbo Chen et al.ACM MM 2023 · 4 citations
- Revisiting Counterfactual Problems in Referring Expression ComprehensionZhihan Yu, Ruifan LiCVPR 2024 · 6 citations
- MuCo: Multi-turn Contrastive Learning for Multimodal Embedding ModelGeonmo Gu, Byeongho Heo, Jaemyung Yu, Jaehui Hwang et al.CVPR 2026 · 2 citations
