VA-MoE: Variables-Adaptive Mixture of Experts for Incremental Weather Forecasting
Hao Chen, Tao Han, Song Guo, Jie Zhang, Yonghan Dong, Yue Yu, Lei Bai
摘要
This paper presents Variables-Adaptive Mixture of Experts (VA-MoE), a novel framework for incremental weather forecasting that dynamically adapts to evolving spatiotemporal patterns in real-time data. Traditional weather prediction models often struggle with exorbitant computational expenditure and the need to continuously update forecasts as new observations arrive. VA-MoE addresses these challenges by leveraging a hybrid architecture of experts, where each expert specializes in capturing distinct sub-patterns of atmospheric variables (e.g., temperature, humidity, wind speed). Moreover, the proposed method employs a variable-adaptive gating mechanism to dynamically select and combine relevant experts based on the input context, enabling efficient knowledge distillation and parameter sharing. This design significantly reduces computational overhead while maintaining high forecast accuracy. Experiments on ERA5 dataset demonstrate that VA-MoE performs comparable against state-of-the-art models in both short-term (e.g., 1-3 days) and long-term (e.g., 5 days) forecasting tasks, with only about 25% of trainable parameters and 50% of the initial training data. Code: https://github.com/chenhao-zju/VAMoE
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- STCast: Adaptive Boundary Alignment for Global and Regional Weather ForecastingHao Chen, Tao Han, Jie Zhang, Song Guo 等CVPR 2026 · 被引用 7 次
- EMFormer: Efficient Multi-Scale Transformer for Accumulative Context Weather Forecastinghao chen, Tao Han, Jie ZHANG, Song Guo 等ICML 2026
它引用的顶会 Paper21
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- ClimaX: A foundation model for weather and climateTung Nguyen, Johannes Brandstetter, Ashish Kapoor, Jayesh K. Gupta 等ICML 2023 · 被引用 426 次
相关 Paper
- EWMoE: An Effective Model for Global Weather Forecasting with Mixture-of-ExpertsLihao Gan, Xin Man, Chenghong Zhang, Jie ShaoAAAI 2025 · 被引用 10 次
- Leveraging Heterogeneous Experts with Advantageous Pattern Memory Learning for Traffic PredictionYueyang Yao, Xingyuan Dai, Yisheng LvICDE 2025 · 被引用 3 次
- Efficient Deweahter Mixture-of-Experts with Uncertainty-Aware Feature-Wise Linear ModulationRongyu Zhang, Yulin Luo, Jiaming Liu, Huanrui Yang 等AAAI 2024 · 被引用 30 次
- Dynamic TMoE: A Drift-Aware Dynamic Mixture of Experts Framework for Non-Stationary Time Series ForecastingJiawen Zhu, Shuhan Liu, Di Weng, Yingcai WuICML 2026
- ARROW: An Adaptive Rollout and Routing Method for Global Weather ForecastingJindong Tian, Yifei Ding, Ronghui Xu, Hao Miao 等ICLR 2026 · 被引用 14 次
