Knowledge-aware Leap-LSTM: Integrating Prior Knowledge into Leap-LSTM towards Faster Long Text Classification
Jinhua Du, Yan Huang, Karo Moilanen
摘要
While widely used in industry, recurrent neural networks (RNNs) are known to have deficiencies in dealing with long sequences (e.g. slow inference, vanishing gradients etc.). Recent research has attempted to accelerate RNN models by developing mechanisms to skip irrelevant words in input. Due to the lack of labelled data, it remains as a challenge to decide which words to skip, especially for low-resource classification tasks. In this paper, we propose Knowledge-Aware Leap-LSTM (KALL), a novel architecture which integrates prior human knowledge (created either manually or automatically) like in-domain keywords, terminologies or lexicons into Leap-LSTM to partially supervise the skipping process. More specifically, we propose a knowledge-oriented cost function for KALL; furthermore, we propose two strategies to integrate the knowledge: (1) the Factored KALL approach involves a keyword indicator as a soft constraint for the skipping process, and (2) the Gated KALL enforces the inclusion of keywords while maintaining a differentiable network in training. Experiments on different public datasets show that our approaches are 1.1x ∼ 2.6x faster than LSTM with better accuracy and 23.6x faster than XLNet in a resourcelimited CPU-only environment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- SkipW: Resource Adaptable RNN with Strict Upper Computational LimitTsiry Mayet, Anne Lambert, Pascal Leguyadec, Françoise Le Bolzer 等ICLR 2021
- A Skip-Connected Evolving Recurrent Neural Network for Data Stream Classification under Label Latency ScenarioMonidipa Das, Mahardhika Pratama, Jie Zhang, Yew-Soon OngAAAI 2020 · 被引用 14 次
- Learning Variational Word Masks to Improve the Interpretability of Neural Text ClassifiersHanjie Chen, Yangfeng JiEMNLP 2020 · 被引用 45 次
- Improving Low-Resource Sequence Labeling with Knowledge Fusion and Contextual Label ExplanationsPeichao Lai, Jiaxin Gan, Feiyang Ye, Wentao Zhang 等EMNLP 2025 · 被引用 1 次
- AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM InferenceZhuomin He, Yizhen Yao, Pengfei Zuo, Bin Gao 等AAAI 2025 · 被引用 13 次
