DialogConv: A Lightweight Fully Convolutional Network for Multi-view Response Selection
Yongkang Liu, Shi Feng, Wei Gao, Daling Wang, Yifei Zhang
摘要
Current end-to-end retrieval-based dialogue systems are mainly based on Recurrent Neural Networks or Transformers with attention mechanisms. Although promising results have been achieved, these models often suffer from slow inference or huge number of parameters. In this paper, we propose a novel lightweight fully convolutional architecture, called DialogConv, for response selection. DialogConv is exclusively built on top of convolution to extract matching features of context and response. Dialogues are modeled in 3D views, where DialogConv performs convolution operations on embedding view, word view and utterance view to capture richer semantic information from multiple contextual views. On the four benchmark datasets, compared with state-of-the-art baselines, Di-alogConv is on average about 8.5× smaller in size, and 79.39× and 10.64× faster on CPU and GPU devices, respectively. At the same time, DialogConv achieves the competitive effectiveness of response selection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 被引用 1,747 次
- Lite Transformer with Long-Short Range AttentionZhanghao Wu, Zhijian Liu, Ji Lin, Yujun Lin 等ICLR 2020 · 被引用 379 次
- MuTual: A Dataset for Multi-Turn Dialogue ReasoningLeyang Cui, Yu Wu, Shujie Liu, Yue Zhang 等ACL 2020 · 被引用 115 次
- A Graph Reasoning Network for Multi-turn Response Selection via Customized Pre-trainingYongkang Liu, Shi Feng, Daling Wang, Kaisong Song 等AAAI 2021 · 被引用 23 次
相关 Paper
- Contextual Fine-to-Coarse Distillation for Coarse-grained Response Selection in Open-Domain ConversationsWei Chen, Yeyun Gong, Can Xu, Huang Hu 等ACL 2022
- Towards Efficient Dialogue Pre-training with Transferable and Interpretable Latent StructureXueliang Zhao, Lemao Liu, Tingchen Fu, Shuming Shi 等EMNLP 2022 · 被引用 3 次
- A Pre-training Strategy for Zero-Resource Response Selection in Knowledge-Grounded ConversationsChongyang Tao, Changyu Chen, Jiazhan Feng, Ji-Rong Wen 等ACL 2021
- DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank UtterancesXiaodong Gu, Kang Min Yoo, Jung-Woo HaAAAI 2021 · 被引用 83 次
- Generating Dialogue Responses from a Semantic Latent SpaceWei-Jen Ko, Avik Ray, Yilin Shen, Hongxia JinEMNLP 2020 · 被引用 4 次
