Learning Reasoning Paths over Semantic Graphs for Video-grounded Dialogues
Hung Le, Nancy F. Chen, Steven C. H. Hoi
摘要
Compared to traditional visual question answering, video-grounded dialogues require additional reasoning over dialogue context to answer questions in a multi-turn setting. Previous approaches to video-grounded dialogues mostly use dialogue context as a simple text input without modelling the inherent information flows at the turn level. In this paper, we propose a novel framework of Reasoning Paths in Dialogue Context (PDC). PDC model discovers information flows among dialogue turns through a semantic graph constructed based on lexical components in each question and answer. PDC model then learns to predict reasoning paths over this semantic graph. Our path prediction model predicts a path from the current turn through past dialogue turns that contain additional visual cues to answer the current question. Our reasoning model sequentially processes both visual and textual information through this reasoning path and the propagated features are used to generate the answer. Our experimental results demonstrate the effectiveness of our method and provide additional insights on how models use semantic dependencies in a dialogue context to retrieve visual cues.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Learning to Retrieve Reasoning Paths over Wikipedia Graph for Question AnsweringAkari Asai, Kazuma Hashimoto, Hannaneh Hajishirzi, Richard Socher 等ICLR 2020 · 被引用 322 次
- Who Did They Respond to? Conversation Structure Modeling Using Masked Hierarchical TransformerHenghui Zhu, Feng Nan, Zhiguo Wang, Ramesh Nallapati 等AAAI 2020 · 被引用 41 次
- Learning to Ask More: Semi-Autoregressive Sequential Question Generation under Dual-Graph InteractionZi Chai, Xiaojun WanACL 2020 · 被引用 26 次
- Detecting Attackable Sentences in ArgumentsYohan Jo, Seojin Bang, Emaad A. Manzoor, Eduard H. Hovy 等EMNLP 2020 · 被引用 25 次
相关 Paper
- DMRM: A Dual-Channel Multi-Hop Reasoning Model for Visual DialogFeilong Chen, Fandong Meng, Jiaming Xu, Peng Li 等AAAI 2020 · 被引用 35 次
- DualVD: An Adaptive Dual Encoding Model for Deep Visual Understanding in Visual DialogueXiaoze Jiang, Jing Yu, Zengchang Qin, Yingying Zhuang 等AAAI 2020 · 被引用 72 次
- DVD: A Diagnostic Dataset for Multi-step Reasoning in Video Grounded DialogueHung Le, Chinnadhurai Sankar, Seungwhan Moon, Ahmad Beirami 等ACL 2021
- BiST: Bi-directional Spatio-Temporal Reasoning for Video-Grounded DialoguesHung Le, Doyen Sahoo, Nancy F. Chen, Steven C. H. HoiEMNLP 2020 · 被引用 30 次
- Inferential Knowledge-Enhanced Integrated Reasoning for Video Question AnsweringJianguo Mao, Wenbin Jiang, Hong Liu, Xiangdong Wang 等AAAI 2023 · 被引用 1 次
