Improving Disfluency Detection by Self-Training a Self-Attentive Model
Paria Jamshid Lou, Mark Johnson
摘要
Self-attentive neural syntactic parsers using contextualized word embeddings (e.g. ELMo or BERT) currently produce state-of-the-art results in joint parsing and disfluency detection in speech transcripts. Since the contextualized word embeddings are pre-trained on a large amount of unlabeled data, using additional unlabeled data to train a neural model might seem redundant. However, we show that self-training -a semi-supervised technique for incorporating unlabeled data -sets a new state-of-the-art for the self-attentive parser on disfluency detection, demonstrating that self-training provides benefits orthogonal to the pre-trained contextualized word representations. We also show that ensembling selftrained parsers provides further gains for disfluency detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Cross2StrA: Unpaired Cross-lingual Image Captioning with Cross-lingual Cross-modal Structure-pivoted AlignmentShengqiong Wu, Hao Fei, Wei Ji, Tat-Seng ChuaACL 2023 · 被引用 42 次
- Planning and Generating Natural and Diverse Disfluent Texts as Augmentation for Disfluency DetectionJingfeng Yang, Diyi Yang, Zhaoran MaEMNLP 2020 · 被引用 13 次
- Combining Self-Training and Self-Supervised Learning for Unsupervised Disfluency DetectionShaolei Wang, Zhongyuan Wang, Wanxiang Che, Ting LiuEMNLP 2020 · 被引用 11 次
- MeetingQA: Extractive Question-Answering on Meeting TranscriptsArchiki Prasad, Trung Bui, Seunghyun Yoon, Hanieh Deilamsalehy 等ACL 2023 · 被引用 4 次
- Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection IncrementalMorteza Rohanian, Julian HoughACL 2021
相关 Paper
- Multi-Task Self-Supervised Learning for Disfluency DetectionShaolei Wang, Wanxiang Che, Qi Liu, Pengda Qin 等AAAI 2020 · 被引用 56 次
- Revisiting Tri-training of Dependency ParsersJoachim Wagner, Jennifer FosterEMNLP 2021
- SenseBERT: Driving Some Sense into BERTYoav Levine, Barak Lenz, Or Dagan, Ori Ram 等ACL 2020 · 被引用 27 次
- Revisiting Self-Training for Neural Sequence GenerationJunxian He, Jiatao Gu, Jiajun Shen, Marc'Aurelio RanzatoICLR 2020 · 被引用 294 次
- Adapting Unsupervised Syntactic Parsing Methodology for Discourse Dependency ParsingLiwen Zhang, Ge Wang, Wenjuan Han, Kewei TuACL 2021
