Automated Concatenation of Embeddings for Structured Prediction
Xinyu Wang, Yong Jiang, Nguyen Bach, Tao Wang, Zhongqiang Huang, Fei Huang, Kewei Tu
摘要
Pretrained contextualized embeddings are powerful word representations for structured prediction tasks. Recent work found that better word representations can be obtained by concatenating different types of embeddings. However, the selection of embeddings to form the best concatenated representation usually varies depending on the task and the collection of candidate embeddings, and the everincreasing number of embedding types makes it a more difficult problem. In this paper, we propose Automated Concatenation of Embeddings (ACE) to automate the process of finding better concatenations of embeddings for structured prediction tasks, based on a formulation inspired by recent progress on neural architecture search. Specifically, a controller alternately samples a concatenation of embeddings, according to its current belief of the effectiveness of individual embedding types in consideration for a task, and updates the belief based on a reward. We follow strategies in reinforcement learning to optimize the parameters of the controller and compute the reward based on the accuracy of a task model, which is fed with the sampled concatenation as input and trained on a task dataset. Empirical results on 6 tasks and 21 datasets show that our approach outperforms strong baselines and achieves state-of-the-art performance with fine-tuned embeddings in all the evaluations. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Is ChatGPT a General-Purpose Natural Language Processing Task Solver?Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen 等EMNLP 2023 · 被引用 449 次
- Packed Levitated Marker for Entity and Relation ExtractionDeming Ye, Yankai Lin, Peng Li, Maosong SunACL 2022 · 被引用 140 次
- Distantly Supervised Named Entity Recognition via Confidence-Based Multi-Class Positive and Unlabeled LearningKang Zhou, Yuepei Li, Qi LiACL 2022 · 被引用 37 次
- Naamapadam: A Large-Scale Named Entity Annotated Data for Indic LanguagesArnav Mhaske, Harshit Kedia, Sumanth Doddapaneni, Mitesh M. Khapra 等ACL 2023 · 被引用 24 次
- PHEE: A Dataset for Pharmacovigilance Event Extraction from TextZhaoyue Sun, Jiazheng Li, Gabriele Pergola, Byron C. Wallace 等EMNLP 2022 · 被引用 17 次
它引用的顶会 Paper8
- LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attentionIkuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda 等EMNLP 2020 · 被引用 562 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- SeqVAT: Virtual Adversarial Training for Semi-Supervised Sequence LabelingLuoxin Chen, Weitong Ruan, Xinyue Liu, Jianhua LuACL 2020 · 被引用 118 次
- Efficient Second-Order TreeCRF for Neural Dependency ParsingYu Zhang, Zhenghua Li, Min ZhangACL 2020 · 被引用 90 次
- Global Greedy Dependency ParsingZuchao Li, Hai Zhao, Kevin ParnowAAAI 2020 · 被引用 34 次
相关 Paper
- Contextual Parameter Generation for Knowledge Graph Link PredictionGeorge Stoica, Otilia Stretcu, Emmanouil Antonios Platanios, Tom M. Mitchell 等AAAI 2020 · 被引用 46 次
- AutoAttend: Automated Attention Representation SearchChaoyu Guan, Xin Wang, Wenwu ZhuICML 2021 · 被引用 46 次
- Structured Prediction as Translation between Augmented Natural LanguagesGiovanni Paolini, Ben Athiwaratkun, Jason Krone, Jie Ma 等ICLR 2021 · 被引用 351 次
- TextNAS: A Neural Architecture Search Space Tailored for Text RepresentationYujing Wang, Yaming Yang, Yiren Chen, Jing Bai 等AAAI 2020 · 被引用 66 次
- Learning Conceptual-Contextual Embeddings for Medical TextXiao Zhang, Dejing Dou, Ji WuAAAI 2020 · 被引用 17 次
