doc2dial: A Goal-Oriented Document-Grounded Dialogue Dataset
Song Feng, Hui Wan, R. Chulaka Gunasekara, Siva Sankalp Patel, Sachindra Joshi, Luis A. Lastras
摘要
We introduce doc2dial, a new dataset of goal-oriented dialogues that are grounded in the associated documents. Inspired by how the authors compose documents for guiding end users, we first construct dialogue flows based on the content elements that corresponds to higher-level relations across text sections as well as lower-level relations between discourse units within a section. Then we present these dialogue flows to crowd contributors to create conversational utterances. The dataset includes over 4500 annotated conversations with an average of 14 turns that are grounded in over 450 documents from four domains. Compared to the prior document-grounded dialogue datasets, this dataset covers a variety of dialogue scenes in information-seeking conversations. For evaluating the versatility of the dataset, we introduce multiple dialogue modeling tasks and present baseline approaches. A9: Would you like to find out whether you are eligible? U10: That's exactly why I contact again! A11: Were there any damages to your clothes that were caused by prosthetic or orthopedic device or your skin medicine? U12: The latter happened.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMsYue Yu, Wei Ping, Zihan Liu, Boxin Wang 等NeurIPS 2024 · 被引用 321 次
- ChatQA: Surpassing GPT-4 on Conversational QA and RAGZihan Liu, Wei Ping, Rajarshi Roy, Peng Xu 等NeurIPS 2024 · 被引用 121 次
- Dialog Inpainting: Turning Documents into DialogsZhuyun Dai, Arun Tejasvi Chaganty, Vincent Y. Zhao, Aida Amini 等ICML 2022 · 被引用 77 次
- InstructRetro: Instruction Tuning post Retrieval-Augmented PretrainingBoxin Wang, Wei Ping, Lawrence McAfee, Peng Xu 等ICML 2024 · 被引用 75 次
- MultiDoc2Dial: Modeling Dialogues Grounded in Multiple DocumentsSong Feng, Siva Sankalp Patel, Hui Wan, Sachindra JoshiEMNLP 2021 · 被引用 42 次
它引用的顶会 Paper1
相关 Paper
- Converse, Focus and Guess - Towards Multi-Document Driven DialogueHan Liu, Caixia Yuan, Xiaojie Wang, Yushu Yang 等AAAI 2021 · 被引用 1 次
- SuperDialseg: A Large-scale Dataset for Supervised Dialogue SegmentationJunfeng Jiang, Chengzhang Dong, Sadao Kurohashi, Akiko AizawaEMNLP 2023 · 被引用 2 次
- Causal Document-Grounded Dialogue Pre-trainingYingxiu Zhao, Bowen Yu, Bowen Li, Haiyang Yu 等EMNLP 2023 · 被引用 2 次
- Interview: Large-scale Modeling of Media Dialog with Discourse Patterns and Knowledge GroundingBodhisattwa Prasad Majumder, Shuyang Li, Jianmo Ni, Julian J. McAuleyEMNLP 2020 · 被引用 12 次
- ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language ModelsJeonghoon Shim, Gyuhyeon Seo, Cheongsu Lim, Yohan JoICLR 2025
