Learning to Contextually Aggregate Multi-Source Supervision for Sequence Labeling
Ouyu Lan, Xiao Huang, Bill Yuchen Lin, He Jiang, Liyuan Liu, Xiang Ren
Abstract
Sequence labeling is a fundamental framework for various natural language processing problems. Its performance is largely influenced by the annotation quality and quantity in supervised learning scenarios, and obtaining ground truth labels is often costly. In many cases, ground truth labels do not exist, but noisy annotations or annotations from different domains are accessible. In this paper, we propose a novel framework Consensus Network (CONNET) that can be trained on annotations from multiple sources (e.g., crowd annotation, cross-domain data...). It learns individual representation for every source and dynamically aggregates source-specific knowledge by a context-aware attention module. Finally, it leads to a model reflecting the agreement (consensus) among multiple sources. We evaluate the proposed framework in two practical settings of multi-source learning: learning with crowd annotations and unsupervised crossdomain model adaptation. Extensive experimental results show that our model achieves significant improvements over existing methods in both settings. We also demonstrate that the method can apply to various tasks and cope with different encoders. 1 * The first two authors contributed equally. 1 Our code can be found at https://github.com/ INK-USC/ConNet .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 956aeb77-0f14-4f6b-a700-87011b0ec542Cited by top-tier papers10
- BOND: BERT-Assisted Open-Domain Named Entity Recognition with Distant SupervisionChen Liang, Yue Yu, Haoming Jiang, Siawpeng Er et al.KDD 2020 · 118 citations
- Navigation Turing Test (NTT): Learning to Evaluate Human-Like NavigationSam Devlin, Raluca Georgescu, Ida Momennejad, Jaroslaw Rzepecki et al.ICML 2021 · 27 citations
- Sparse Conditional Hidden Markov Model for Weakly Supervised Named Entity RecognitionYinghao Li, Le Song, Chao ZhangKDD 2022 · 15 citations
- Understanding Programmatic Weak Supervision via Source-aware Influence FunctionJieyu Zhang, Haonan Wang, Cheng-Yu Hsieh, Alexander J. RatnerNeurIPS 2022 · 13 citations
- Addressing NER Annotation Noises with Uncertainty-Guided Tree-Structured CRFsJian Liu, Weichang Liu, Yufeng Chen, Jinan Xu et al.EMNLP 2023 · 3 citations
Builds on1
Related papers
- Crowdsourcing Learning as Domain Adaptation: A Case Study on Named Entity RecognitionXin Zhang, Guangwei Xu, Yueheng Sun, Meishan Zhang et al.ACL 2021
- Spotting the Unseen: Reciprocal Consensus Network Guided by Visual ArchetypesWenbo Hu, Hongjian Zhan, Xinchen Ma, Yue Lu et al.AAAI 2024 · 2 citations
- Unbiased Multi-Label Learning from Crowdsourced AnnotationsMingxuan Xia, Zenan Huang, Runze Wu, Gengyu Lyu et al.ICML 2024 · 1 citation
- Universal Domain Adaptive Network Embedding for Node ClassificationJushuo Chen, Feifei Dai, Xiaoyan Gu, Jiang Zhou et al.ACM MM 2023 · 4 citations
- Learning from Crowds by Modeling Common ConfusionsZhendong Chu, Jing Ma, Hongning WangAAAI 2021 · 60 citations
