A Lightweight Sparse Interaction Network for Time Series Forecasting
Xu Zhang, Qitong Wang, Peng Wang, Wei Wang
Abstract
Recent work shows that linear models can outperform several transformer models in long-term time-series forecasting (TSF). However, instead of explicitly performing temporal interaction through self-attention, linear models implicitly perform it based on stacked MLP structures, which may be insufficient in capturing the complex temporal dependencies and their performance still has potential for improvement. To this end, we propose a Lightweight Sparse Interaction Network (LSINet) for TSF task. Inspired by the sparsity of selfattention, we propose a Multihead Sparse Interaction Mechanism (MSIM). Different from self-attention, MSIM learns the important connections between time steps through sparsityinduced Bernoulli distribution to capture temporal dependencies for TSF. The sparsity is ensured by the proposed self-adaptive regularization loss. Moreover, we observe the shareability of temporal interactions and propose to perform Shared Interaction Learning (SIL) for MSIM to further enhance efficiency and improve convergence. LSINet is a linear model comprising only MLP structures with low overhead and equipped with explicit temporal interaction mechanisms. Extensive experiments on public datasets show that LSINet achieves both higher accuracy and better efficiency than advanced linear models and transformer models in TSF tasks. The code is available at the link https://github.com/Meteor- Stars/LSINet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b303a02d-ca40-4e54-ba9d-2c8f06eb13c6Cited by top-tier papers5
- Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?Xu Zhang, Peng Wang, Wei WangICLR 2026 · 2 citations
- Diff-MN: Diffusion Parameterized MoE-NCDE for Continuous Time Series Generation with Irregular ObservationsXu Zhang, Junwei Deng, Chang Xu, Hao Li et al.ICML 2026 · 2 citations
- Amortized Predictability-aware Training Framework for Time Series Forecasting and ClassificationXu Zhang, Peng Wang, Yichen Li, Wei WangWWW 2026
- SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series ForecastingXu Zhang, Qitong Wang, Peng Wang, Wei WangWWW 2026
- Revisiting Network Inertia: Dynamic Inertia Inhibition Coupled Multidimensional Periodicity for Infrared and Visible Image FusionYufeng Chen, Yuan Sun, Hao Pan, Xujian Zhao et al.AAAI 2026
Builds on10
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang et al.ICML 2022 · 2,912 citations
- One Fits All: Power General Time Series Analysis by Pretrained LMTian Zhou, Peisong Niu, Xue Wang, Liang Sun et al.NeurIPS 2023 · 1,178 citations
Related papers
- Sparse-Scale Transformer with Bidirectional Awareness for Time Series ForecastingYing Liu, Bo Liu, Sheng Huang, Gang Luo et al.AAAI 2026
- Are Self-Attentions Effective for Time Series Forecasting?Dongbin Kim, Jinseong Park, Jaewook Lee, Hoki KimNeurIPS 2024 · 48 citations
- SAMformer: Unlocking the Potential of Transformers in Time Series Forecasting with Sharpness-Aware Minimization and Channel-Wise AttentionRomain Ilbert, Ambroise Odonnat, Vasilii Feofanov, Aladin Virmaux et al.ICML 2024 · 62 citations
- Linear Transformers as VAR Models: Aligning Autoregressive Attention Mechanisms with Autoregressive ForecastingJiecheng Lu, Shihao YangICML 2025
- Unlocking the Power of Patch: Patch-Based MLP for Long-Term Time Series ForecastingPeiwang Tang, Weitai ZhangAAAI 2025 · 42 citations
