Knowledge-aware Leap-LSTM: Integrating Prior Knowledge into Leap-LSTM towards Faster Long Text Classification
Jinhua Du, Yan Huang, Karo Moilanen
Abstract
While widely used in industry, recurrent neural networks (RNNs) are known to have deficiencies in dealing with long sequences (e.g. slow inference, vanishing gradients etc.). Recent research has attempted to accelerate RNN models by developing mechanisms to skip irrelevant words in input. Due to the lack of labelled data, it remains as a challenge to decide which words to skip, especially for low-resource classification tasks. In this paper, we propose Knowledge-Aware Leap-LSTM (KALL), a novel architecture which integrates prior human knowledge (created either manually or automatically) like in-domain keywords, terminologies or lexicons into Leap-LSTM to partially supervise the skipping process. More specifically, we propose a knowledge-oriented cost function for KALL; furthermore, we propose two strategies to integrate the knowledge: (1) the Factored KALL approach involves a keyword indicator as a soft constraint for the skipping process, and (2) the Gated KALL enforces the inclusion of keywords while maintaining a differentiable network in training. Experiments on different public datasets show that our approaches are 1.1x ∼ 2.6x faster than LSTM with better accuracy and 23.6x faster than XLNet in a resourcelimited CPU-only environment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- SkipW: Resource Adaptable RNN with Strict Upper Computational LimitTsiry Mayet, Anne Lambert, Pascal Leguyadec, Françoise Le Bolzer et al.ICLR 2021
- A Skip-Connected Evolving Recurrent Neural Network for Data Stream Classification under Label Latency ScenarioMonidipa Das, Mahardhika Pratama, Jie Zhang, Yew-Soon OngAAAI 2020 · 14 citations
- Learning Variational Word Masks to Improve the Interpretability of Neural Text ClassifiersHanjie Chen, Yangfeng JiEMNLP 2020 · 45 citations
- Improving Low-Resource Sequence Labeling with Knowledge Fusion and Contextual Label ExplanationsPeichao Lai, Jiaxin Gan, Feiyang Ye, Wentao Zhang et al.EMNLP 2025 · 1 citation
- AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM InferenceZhuomin He, Yizhen Yao, Pengfei Zuo, Bin Gao et al.AAAI 2025 · 13 citations
