Multitask Semi-Supervised Learning for Class-Imbalanced Discourse Classification
Alexander Spangher, Jonathan May, Sz-Rung Shiang, Lingjia Deng
Abstract
As labeling schemas evolve over time, small differences can render datasets following older schemas unusable. This prevents researchers from building on top of previous annotation work and results in the existence, in discourse learning in particular, of many small classimbalanced datasets. In this work, we show that a multitask learning approach can combine discourse datasets from similar and diverse domains to improve discourse classification. We show an improvement of 4.9% Micro F1-score over current state-of-the-art benchmarks on the NewsDiscourse dataset, one of the largest discourse datasets recently published, due in part to label correlations across tasks, which improve performance for underrepresented classes. We also offer an extensive review of additional techniques proposed to address resource-poor problems in NLP, and show that none of these approaches can improve classification accuracy in our setting 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Are Large Language Models Capable of Generating Human-Level Narratives?Yufei Tian, Tenghao Huang, Miri Liu, Derek Jiang et al.EMNLP 2024 · 22 citations
- Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMsAlexander Spangher, Nanyun Peng, Sebastian Gehrmann, Mark DredzeEMNLP 2024 · 4 citations
Builds on5
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- Dice Loss for Data-imbalanced NLP TasksXiaoya Li, Xiaofei Sun, Yuxian Meng, Junjun Liang et al.ACL 2020 · 575 citations
- On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong BaselinesMarius Mosbach, Maksym Andriushchenko, Dietrich KlakowICLR 2021 · 448 citations
- MixText: Linguistically-Informed Interpolation of Hidden Space for Semi-Supervised Text ClassificationJiaao Chen, Zichao Yang, Diyi YangACL 2020 · 340 citations
- Discourse as a Function of Event: Profiling Discourse Structure in News Articles around the Main EventPrafulla Kumar Choubey, Aaron Lee, Ruihong Huang, Lu WangACL 2020 · 54 citations
Related papers
- TED-CDB: A Large-Scale Chinese Discourse Relation Dataset on TED TalksWanqiu Long, Bonnie Webber, Deyi XiongEMNLP 2020 · 11 citations
- Multi-Task Self-Supervised Learning for Disfluency DetectionShaolei Wang, Wanxiang Che, Qi Liu, Pengda Qin et al.AAAI 2020 · 56 citations
- GDTB: Genre Diverse Data for English Shallow Discourse Parsing across Modalities, Text Types, and DomainsYang Janet Liu, Tatsuya Aoyama, Wesley Scivetti, Yilun Zhu et al.EMNLP 2024
- Unsupervised Learning of Discourse Structures using a Tree AutoencoderPatrick Huber, Giuseppe CareniniAAAI 2021 · 4 citations
- Meta Self-training for Few-shot Neural Sequence LabelingYaqing Wang, Subhabrata Mukherjee, Haoda Chu, Yuancheng Tu et al.KDD 2021 · 56 citations
