Poseidon: Efficient Foundation Models for PDEs
Maximilian Herde, Bogdan Raonic, Tobias Rohner, Roger Käppeli, Roberto Molinaro, Emmanuel de Bézenac, Siddhartha Mishra
Abstract
We introduce Poseidon, a foundation model for learning the solution operators of PDEs. It is based on a multiscale operator transformer, with time-conditioned layer norms that enable continuous-in-time evaluations. A novel training strategy leveraging the semi-group property of time-dependent PDEs to allow for significant scaling-up of the training data is also proposed. Poseidon is pretrained on a diverse, large scale dataset for the governing equations of fluid dynamics. It is then evaluated on a suite of 15 challenging downstream tasks that include a wide variety of PDE types and operators. We show that Poseidon exhibits excellent performance across the board by outperforming baselines significantly, both in terms of sample efficiency and accuracy. Poseidon also generalizes very well to new physics that is not seen during pretraining. Moreover, Poseidon scales with respect to model and data size, both for pretraining and for downstream tasks. Taken together, our results showcase the surprising ability of Poseidon to learn effective representations from a very small set of PDEs during pretraining in order to generalize well to unseen and unrelated PDEs downstream, demonstrating its potential as an effective, general purpose PDE foundation model. Finally, the Poseidon model as well as underlying pretraining and downstream datasets are open sourced, with code being available at https://github.com/camlab-ethz/poseidon and pretrained models and datasets at https://huggingface.co/camlab-ethz.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 09fad57e-2c7b-450f-b090-187e804bc35eCited by top-tier papers33
- Geometry Aware Operator Transformer as an efficient and accurate neural surrogate for PDEs on arbitrary domainsShizheng Wen, Arsh Kumbhat, Levi E. Lingsch, Sepehr Mousavi et al.NeurIPS 2025 · 73 citations
- Physics vs Distributions: Pareto Optimal Flow Matching with Physics ConstraintsGiacomo Baldan, Qiang Liu, Alberto Guardone, Nils ThuereyICLR 2026 · 31 citations
- RIGNO: A Graph-based Framework For Robust And Accurate Operator Learning For PDEs On Arbitrary DomainsSepehr Mousavi, Shizheng Wen, Levi E. Lingsch, Maximilian Herde et al.NeurIPS 2025 · 31 citations
- Lost in Latent Space: An Empirical Study of Latent Diffusion Models for Physics EmulationFrançois Rozet, Ruben Ohana, Michael McCabe, Gilles Louppe et al.NeurIPS 2025 · 23 citations
- RealPDEBench: A Benchmark for Complex Physical Systems with Real-World DataPeiyan Hu, Haodong Feng, Hongyuan Liu, Tongtong Yan et al.ICLR 2026 · 17 citations
Builds on20
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
Related papers
- DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-TrainingZhongkai Hao, Chang Su, Songming Liu, Julius Berner et al.ICML 2024 · 107 citations
- PDE-Transformer: Efficient and Versatile Transformers for Physics SimulationsBenjamin J. Holzschuh, Qiang Liu, Georg Kohl, Nils ThuereyICML 2025
- Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE SolversHang Zhou, Yuezhou Ma, Haixu Wu, Haowen Wang et al.ICML 2025
- GNOT: A General Neural Operator Transformer for Operator LearningZhongkai Hao, Zhengyi Wang, Hang Su, Chengyang Ying et al.ICML 2023 · 375 citations
- Multiple Physics Pretraining for Spatiotemporal Surrogate ModelsMichael McCabe, Bruno Régaldo-Saint Blancard, Liam Holden Parker, Ruben Ohana et al.NeurIPS 2024 · 97 citations
