A Sequence-to-Sequence Approach with Mixed Pointers to Topic Segmentation and Segment Labeling
Jinxiong Xia, Houfeng Wang
摘要
Topic segmentation is the process of dividing a text into semantically coherent segments, and segment labeling involves assigning a topic label to each of these segments. Previous work on this task has included the use of sequence labeling, segment-extraction, and generative models. While these methods have yielded impressive results, existing generative models have struggled to accurately generate strings of segment boundaries, limiting their competitiveness in this area. In this paper, we present a novel Sequence-to-Sequence approach with Mixed Pointers (Seq2Seq-MP). Seq2Seq-MP employs an encoder-decoder architecture with the pointer mechanism to generate both segment boundaries and topics, which allows for a more robust performance than string-generation models and can handle long-range dependencies better than sequence labeling and segment-extraction models. Additionally, we introduce the pairwise type encoding and type-aware relative position encoding to improve the fusion of type and position information, enhancing the interactions between sentences and topics in the encoder and decoder. Our experiments on public datasets show that Seq2Seq-MP outperforms the current state-of-the-art, with up to 2.9% and 4.0% improvements in Pk and F1, respectively.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- SegFormer: A Topic Segmentation Model with Controllable Range of AttentionHaitao Bai, Pinghui Wang, Ruofei Zhang, Zhou SuAAAI 2023 · 被引用 18 次
- An Autoregressive Text-to-Graph Framework for Joint Entity and Relation ExtractionUrchade Zaratiana, Nadi Tomeh, Pierre Holat, Thierry CharnoisAAAI 2024 · 被引用 39 次
- Improving Long Document Topic Segmentation Models With Enhanced Coherence ModelingHai Yu, Chong Deng, Qinglin Zhang, Jiaqing Liu 等EMNLP 2023 · 被引用 7 次
- Rethinking Boundaries: End-To-End Recognition of Discontinuous Mentions with Pointer NetworksHao Fei, Donghong Ji, Bobo Li, Yijiang Liu 等AAAI 2021 · 被引用 82 次
- A Joint Model for Document Segmentation and Segment LabelingJoe Barrow, Rajiv Jain, Vlad I. Morariu, Varun Manjunatha 等ACL 2020 · 被引用 47 次
