Deep or Simple Models for Semantic Tagging? It Depends on your Data
Jinfeng Li, Yuliang Li, Xiaolan Wang, Wang-Chiew Tan
摘要
Semantic tagging, which has extensive applications in text mining, predicts whether a given piece of text conveys the meaning of a given semantic tag. The problem of semantic tagging is largely solved with supervised learning and today, deep learning models are widely perceived to be better for semantic tagging. However, there is no comprehensive study supporting the popular belief. Practitioners often have to train different types of models for each semantic tagging task to identify the best model. This process is both expensive and inefficient. We embark on a systematic study to investigate the following question: Are deep models the best performing model for all semantic tagging tasks? To answer this question, we compare deep models against "simple models" over datasets with varying characteristics. Specifically, we select three prevalent deep models (i.e. CNN, LSTM, and BERT) and two simple models (i.e. LR and SVM), and compare their performance on the semantic tagging task over 21 datasets. Results show that the size, the label ratio, and the label cleanliness of a dataset significantly impact the quality of semantic tagging. Simple models achieve similar tagging quality to deep models on large datasets, but the runtime of simple models is much shorter. Moreover, simple models can achieve better tagging quality than deep models when targeting datasets show worse label cleanliness and/or more severe imbalance. Based on these findings, our study can systematically guide practitioners in selecting the right learning model for their semantic tagging task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Active Learning for BERT: An Empirical StudyLiat Ein-Dor, Alon Halfon, Ariel Gera, Eyal Shnarch 等EMNLP 2020 · 被引用 144 次
- Making the Relation Matters: Relation of Relation Learning Network for Sentence Semantic MatchingKun Zhang, Le Wu, Guangyi Lv, Meng Wang 等AAAI 2021 · 被引用 25 次
- Label Confusion Learning to Enhance Text Classification ModelsBiyang Guo, Songqiao Han, Xiao Han, Hailiang Huang 等AAAI 2021 · 被引用 79 次
- Pre-trained Embeddings for Entity Resolution: An Experimental AnalysisAlexandros Zeakis, George Papadakis, Dimitrios Skoutas, Manolis KoubarakisVLDB 2023 · 被引用 63 次
- FLiText: A Faster and Lighter Semi-Supervised Text Classification with Convolution NetworksChen Liu, Mengchao Zhang, Zhibing Fu, Panpan Hou 等EMNLP 2021 · 被引用 14 次
