MixLinear: Extreme Low Resource Multivariate Time Series Forecasting with 0.1K Parameters
Aitian Ma, Dongsheng Luo, Mo Sha
Abstract
Recently, there has been a growing interest in Long-term Time Series Forecasting (LTSF), which involves predicting long-term future values by analyzing a large amount of historical time-series data to identify patterns and trends. Significant challenges exist in LTSF due to its complex temporal dependencies and high computational demands. Although Transformer-based models offer high forecasting accuracy, they are often too compute-intensive to be deployed on devices with hardware constraints. In this paper, we propose MixLinear, which synergistically combines segment-based trend extraction in the time domain with adaptive low-rank spectral filtering in the frequency domain. Our approach exploits the complementary structural sparsity of time series: local temporal patterns are efficiently captured through mathematically linear transformations that separate intra-segment and inter-segment correlations, while global trends are compressed into an ultralow-dimensional frequency latent space through learnable rank-constrained filters. By reducing the parameter scale of a downsampled n-length input/output onelayer linear model from O(n 2 ) to O(n), MixLinear achieves efficient computation without sacrificing accuracy. Extensive evaluations show that MixLinear achieves forecasting performance comparable to existing models with significantly fewer parameters (0.1K), which makes it well-suited for deployment on devices with limited computational capacity. Recent research has started to process local and global components differently. One approach, found in models like DeepGate (Park et al., 2022) , decomposes the time series first. However, such a method *
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Efficient Source-Free Time-Series Adaptation via Parameter Subspace DisentanglementGaurav Patel, Christopher Michael Sandino, Behrooz Mahasseni, Ellen L. Zippi et al.ICLR 2025
- TimeSeed: Effective Time Series Forecasting with Sparse Endogenous VariablesZhaowang Wu, Kaixin Deng, Hua YanICML 2026
Builds on20
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
Related papers
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang et al.ICML 2022 · 2,912 citations
- SparseTSF: Modeling Long-term Time Series Forecasting with 1k ParametersShengsheng Lin, Weiwei Lin, Wentai Wu, Haojun Chen et al.ICML 2024 · 155 citations
- TimeBase: The Power of Minimalism in Efficient Long-term Time Series ForecastingQihe Huang, Zhengyang Zhou, Kuo Yang, Zhongchao Yi et al.ICML 2025
- WaveletMixer: A Multi-Resolution Wavelets Based MLP-Mixer for Multivariate Long-Term Time Series ForecastingZichi Zhang, Tuan Dung Pham, Yimeng An, Ngoc Phu Doan et al.AAAI 2025 · 3 citations
- HMformer: Unleashing Transformer's Potential for Time Series Forecasting via Hierarchical Multi-Scale ModelingRenjun Huang, Han Xiao, Bingqing Li, Baili Zhang et al.AAAI 2026
