Guiding Attention in Sequence-to-Sequence Models for Dialogue Act Prediction
Pierre Colombo, Emile Chapuis, Matteo Manica, Emmanuel Vignon, Giovanna Varni, Chloé Clavel
Abstract
The task of predicting dialog acts (DA) based on conversational dialog is a key component in the development of conversational agents. Accurately predicting DAs requires a precise modeling of both the conversation and the global tag dependencies. We leverage seq2seq approaches widely adopted in Neural Machine Translation (NMT) to improve the modelling of tag sequentiality. Seq2seq models are known to learn complex global dependencies while currently proposed approaches using linear conditional random fields (CRF) only model local tag dependencies. In this work, we introduce a seq2seq model tailored for DA classification using: a hierarchical encoder, a novel guided attention mechanism and beam search applied to both training and inference. Compared to the state of the art our model does not require handcrafted features and is trained end-to-end. Furthermore, the proposed approach achieves an unmatched accuracy score of 85% on SwDA, and state-of-the-art accuracy score of 91.6% on MRDA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14048bc3-a003-4262-86ca-a6543d2da666Cited by top-tier papers7
- Heavy-tailed Representations, Text Polarity Classification & Data AugmentationHamid Jalalzai, Pierre Colombo, Chloé Clavel, Éric Gaussier et al.NeurIPS 2020 · 33 citations
- Automatic Text Evaluation through the Lens of Wasserstein BarycentersPierre Colombo, Guillaume Staerman, Chloé Clavel, Pablo PiantanidaEMNLP 2021 · 21 citations
- What are the best Systems? New Perspectives on NLP BenchmarkingPierre Colombo, Nathan Noiry, Ekhine Irurozki, Stéphan ClémençonNeurIPS 2022 · 20 citations
- Code-switched inspired losses for spoken dialog representationsPierre Colombo, Emile Chapuis, Matthieu Labeau, Chloé ClavelEMNLP 2021 · 6 citations
- Learning Disentangled Textual Representations via Statistical Measures of SimilarityPierre Colombo, Guillaume Staerman, Nathan Noiry, Pablo PiantanidaACL 2022
Related papers
- Variational Hierarchical Dialog Autoencoder for Dialog State Tracking Data AugmentationKang Min Yoo, Hanbit Lee, Franck Dernoncourt, Trung Bui et al.EMNLP 2020
- Structured and Natural Responses Co-generation for Conversational SearchChenchen Ye, Lizi Liao, Fuli Feng, Wei Ji et al.SIGIR 2022 · 21 citations
- A Label Dependence-Aware Sequence Generation Model for Multi-Level Implicit Discourse Relation RecognitionChangxing Wu, Liuwen Cao, Yubin Ge, Yang Liu et al.AAAI 2022 · 38 citations
- Improving Knowledge-Aware Dialogue Generation via Knowledge Base Question AnsweringJian Wang, Junhao Liu, Wei Bi, Xiaojiang Liu et al.AAAI 2020 · 52 citations
- Towards Making the Most of Dialogue Characteristics for Neural Chat TranslationYunlong Liang, Chulun Zhou, Fandong Meng, Jinan Xu et al.EMNLP 2021 · 12 citations
