Do LLMs Understand Dialogues? A Case Study on Dialogue Acts
Ayesha Qamar, Jonathan Tong, Ruihong Huang
Abstract
Recent advancements in NLP, largely driven by Large Language Models (LLMs), have significantly improved performance on an array of tasks. However, Dialogue Act (DA) classification remains challenging, particularly in the fine-grained 50-class, multiparty setting. This paper investigates the root causes of LLMs’ poor performance in DA classification through a linguistically motivated analysis. We identify three key pre-tasks essential for accurate DA prediction: Turn Management , Communica-tive Function Identification , and Dialogue Structure Prediction . Our experiments reveal that LLMs struggle with these fundamental tasks, often failing to outperform simple rule-based baselines. Additionally, we establish a strong empirical correlation between errors in these pre-tasks and DA classification failures. A human study further highlights the significant gap between LLM and human-level dialogue understanding. These findings indicate that LLMs’ shortcomings in dialogue comprehension hinder their ability to accurately predict DAs, highlighting the need for improved dialogue-aware training approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f98b72f6-57e2-4b7a-b1f4-bdc63bb50537Builds on5
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo et al.NeurIPS 2022 · 8,168 citations
- Can Large Language Models be Good Emotional Supporter? Mitigating Preference Bias on Emotional Support ConversationDongjin Kang, Sunghwan Kim, Taeyoon Kwon, Seungjun Moon et al.ACL 2024 · 14 citations
- Zero-Shot Cross-Domain Dialogue State Tracking via Dual Low-Rank AdaptationXiang Luo, Zhiwen Tang, Jin Wang, Xuejie ZhangACL 2024 · 5 citations
- ESCoT: Towards Interpretable Emotional Support Dialogue SystemsTenggan Zhang, Xinjie Zhang, Jinming Zhao, Li Zhou et al.ACL 2024
- Towards Emotional Support Dialog SystemsSiyang Liu, Chujie Zheng, Orianna Demasi, Sahand Sabour et al.ACL 2021
Related papers
- MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn DialoguesGe Bai, Jie Liu, Xingyuan Bu, Yancheng He et al.ACL 2024 · 35 citations
- Evaluating the Effectiveness of Large Language Models in Establishing Conversational GroundingBiswesh Mohapatra, Manav Nitin Kapadnis, Laurent Romary, Justine CassellEMNLP 2024 · 1 citation
- Probing LLMs for Multilingual Discourse Generalization Through a Unified Label SetFlorian Eichin, Yang Janet Liu, Barbara Plank, Michael A. HedderichACL 2025
- Masking Orchestration: Multi-Task Pretraining for Multi-Role Dialogue Representation LearningTianyi Wang, Yating Zhang, Xiaozhong Liu, Changlong Sun et al.AAAI 2020 · 8 citations
- Probing Task-Oriented Dialogue Representation from Language ModelsChien-Sheng Wu, Caiming XiongEMNLP 2020 · 20 citations
