Chaos Meets Attention: Transformers for Large-Scale Dynamical Prediction
Yi He, Yiming Yang, Xiaoyuan Cheng, Hai Wang, Xiao Xue, Boli Chen, Yukun Hu
摘要
Generating long-term trajectories of dissipative chaotic systems autoregressively is a highly challenging task. The inherent positive Lyapunov exponents amplify prediction errors over time. Many chaotic systems possess a crucial property -ergodicity on their attractors, which makes long-term prediction possible. State-of-the-art methods address ergodicity by preserving statistical properties using optimal transport techniques. However, these methods face scalability challenges due to the curse of dimensionality when matching distributions. To overcome this bottleneck, we propose a scalable transformerbased framework capable of stably generating long-term high-dimensional and high-resolution chaotic dynamics while preserving ergodicity. Our method is grounded in a physical perspective, revisiting the Von Neumann mean ergodic theorem to ensure the preservation of long-term statistics in the L 2 space. We introduce novel modifications to the attention mechanism, making the transformer architecture well-suited for learning large-scale chaotic systems. Compared to operator-based and transformer-based methods, our model achieves better performances across five metrics, from short-term prediction accuracy to long-term statistics. In addition to our methodological contributions, we introduce new chaotic system benchmarks: a machine learning dataset of 140k snapshots of turbulent channel flow and a processed high-dimensional Kolmogorov Flow dataset, along with various evaluation metrics for both short-and long-term performances. Both are well-suited for machine learning research on chaotic systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Context parroting: A simple but tough-to-beat baseline for foundation models in scientific machine learningYuanzhao Zhang, William GilpinICLR 2026 · 被引用 16 次
- ChaosNexus: A Foundation Model for ODE-based Chaotic System Forecasting with Hierarchical Multi-scale AwarenessChang Liu, Bohao Zhao, Jingtao Ding, Yong LiICML 2026 · 被引用 1 次
- MMPD-Bench: Bridging Multimodal Fission with Multi-Polarimetric Modalities DecompositionYi He, Zimo Zhao, Yiming Yang, Xiaoyuan Cheng 等ICML 2026
- Tensor-Var: Efficient Four-Dimensional Variational Data AssimilationYiming Yang, Xiaoyuan Cheng, Daniel Giles, Sibo Cheng 等ICML 2025
- Multi-Scale Wavelet Transformers for Operator Learning of Dynamical SystemsXuesong Wang, Michael Groom, Rafael Oliveira, He Zhao 等ICML 2026
它引用的顶会 Paper13
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- GNOT: A General Neural Operator Transformer for Operator LearningZhongkai Hao, Zhengyi Wang, Hang Su, Chengyang Ying 等ICML 2023 · 被引用 375 次
- Multiwavelet-based Operator Learning for Differential EquationsGaurav Gupta, Xiongye Xiao, Paul BogdanNeurIPS 2021 · 被引用 355 次
- Scalable Transformer for PDE Surrogate ModelingZijie Li, Dule Shu, Amir Barati FarimaniNeurIPS 2023 · 被引用 188 次
相关 Paper
- Learning Chaotic Dynamics in Dissipative SystemsZongyi Li, Miguel Liu-Schiaffini, Nikola B. Kovachki, Kamyar Azizzadenesheli 等NeurIPS 2022 · 被引用 62 次
- Learning Chaos In A Linear WayXiaoyuan Cheng, Yi He, Yiming Yang, Xiao Xue 等ICLR 2025
- DySLIM: Dynamics Stable Learning by Invariant Measure for Chaotic SystemsYair Schiff, Zhong Yi Wan, Jeffrey B. Parker, Stephan Hoyer 等ICML 2024 · 被引用 30 次
- Training neural operators to preserve invariant measures of chaotic attractorsRuoxi Jiang, Peter Y. Lu, Elena Orlova, Rebecca WillettNeurIPS 2023 · 被引用 59 次
- When are dynamical systems learned from time series data statistically accurate?Jeongjin Park, Nicole Yang, Nisha ChandramoorthyNeurIPS 2024 · 被引用 17 次
