Multiple Physics Pretraining for Spatiotemporal Surrogate Models
Michael McCabe, Bruno Régaldo-Saint Blancard, Liam Holden Parker, Ruben Ohana, Miles D. Cranmer, Alberto Bietti, Michael Eickenberg, Siavash Golkar, Géraud Krawezik, François Lanusse, Mariel Pettee, Tiberiu Tesileanu
摘要
We introduce multiple physics pretraining (MPP), an autoregressive task-agnostic pretraining approach for physical surrogate modeling of spatiotemporal systems with transformers.In MPP, rather than training one model on a specific physical system, we train a backbone model to predict the dynamics of multiple heterogeneous physical systems simultaneously in order to learn features that are broadly useful across systems and facilitate transfer.In order to learn effectively in this setting, we introduce a shared embedding and normalization strategy that projects the fields of multiple systems into a shared embedding space.We validate the efficacy of our approach on both pretraining and downstream tasks over a broad fluid mechanics-oriented benchmark.We show that a single MPP-pretrained transformer is able to match or outperform task-specific baselines on all pretraining sub-tasks without the need for finetuning.For downstream tasks, we demonstrate that finetuning MPP-trained models results in more accurate predictions across multiple time-steps on systems with previously unseen physical components or higher dimensional systems compared to training from scratch or finetuning pretrained video foundation models.We open-source our code and model weights trained at multiple scales for reproducibility.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- RIGNO: A Graph-based Framework For Robust And Accurate Operator Learning For PDEs On Arbitrary DomainsSepehr Mousavi, Shizheng Wen, Levi E. Lingsch, Maximilian Herde 等NeurIPS 2025 · 被引用 31 次
- Lost in Latent Space: An Empirical Study of Latent Diffusion Models for Physics EmulationFrançois Rozet, Ruben Ohana, Michael McCabe, Gilles Louppe 等NeurIPS 2025 · 被引用 23 次
- Context parroting: A simple but tough-to-beat baseline for foundation models in scientific machine learningYuanzhao Zhang, William GilpinICLR 2026 · 被引用 16 次
- Panda: A pretrained forecast model for chaotic dynamicsJeffrey B. Lai, Anthony Bao, William GilpinICLR 2026 · 被引用 13 次
- Neural Emulator Superiority: When Machine Learning for PDEs Surpasses its Training DataFelix Koehler, Nils ThuereyNeurIPS 2025 · 被引用 8 次
它引用的顶会 Paper21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
相关 Paper
- PDE-Transformer: Efficient and Versatile Transformers for Physics SimulationsBenjamin J. Holzschuh, Qiang Liu, Georg Kohl, Nils ThuereyICML 2025
- P3D: Highly Scalable 3D Neural Surrogates for Physics Simulations with Global ContextBenjamin Holzschuh, Georg Kohl, Florian Redinger, Nils ThuereyICLR 2026 · 被引用 4 次
- Poseidon: Efficient Foundation Models for PDEsMaximilian Herde, Bogdan Raonic, Tobias Rohner, Roger Käppeli 等NeurIPS 2024 · 被引用 235 次
- Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained TransformersLirui Wang, Xinlei Chen, Jialiang Zhao, Kaiming HeNeurIPS 2024 · 被引用 208 次
- UniSTD: Towards Unified Spatio-Temporal Learning across Diverse DisciplinesChen Tang, Xinzhu Ma, Encheng Su, Xiufeng Song 等CVPR 2025
