SC2023Top-tier venue
High Throughput Training of Deep Surrogates from Large Ensemble Runs
Lucas Thibaut Meyer, Marc Schouler, Robert Alexander Caulk, Alejandro Ribés, Bruno Raffin
Abstract
Recent years have seen a surge in deep learning approaches to accelerate numerical solvers, which provide faithful but computationally intensive simulations of the physical world. These deep surrogates are generally trained in a supervised manner from limited amounts of data slowly generated by the same solver they intend to accelerate. We propose an open-source framework that enables the online training of these models from a large ensemble run of simulations. It leverages multiple levels of parallelism to generate rich datasets. The framework avoids I/O bottlenecks and storage issues by directly streaming the generated data. A training reservoir mitigates the inherent bias of streaming while maximizing GPU throughput. Experiment on training a fully connected network as a surrogate for the heat equation shows the proposed approach enables training on 8TB of data in 2 hours with an accuracy improved by 47% and a batch throughput multiplied by 13 compared to a traditional offline procedure.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- Characterizing possible failure modes in physics-informed neural networksAditi S. Krishnapriyan, Amir Gholami, Shandian Zhe, Robert M. Kirby et al.NeurIPS 2021 · 1,421 citations
- Learning Mesh-Based Simulation with Graph NetworksTobias Pfaff, Meire Fortunato, Alvaro Sanchez-Gonzalez, Peter W. BattagliaICLR 2021 · 1,175 citations
- Message Passing Neural PDE SolversJohannes Brandstetter, Daniel E. Worrall, Max WellingICLR 2022 · 410 citations
Related papers
- Training Deep Surrogate Models with Large Scale Online LearningLucas Thibaut Meyer, Marc Schouler, Robert Alexander Caulk, Alejandro Ribés et al.ICML 2023 · 10 citations
- Transolver-3: Scaling Up Transformer Solvers to Industrial-Scale GeometriesHang Zhou, Haixu Wu, Haonan Shangguan, Yuezhou Ma et al.ICML 2026
- Distributed multigrid neural solvers on megavoxel domainsAditya Balu, Sergio Botelho, Biswajit Khara, Vinay Rao et al.SC 2021 · 7 citations
- Learning Incompressible Fluid Dynamics from Scratch - Towards Fast, Differentiable Fluid Models that GeneralizeNils Wandel, Michael Weinmann, Reinhard KleinICLR 2021 · 82 citations
- HelioX: A GPU-Native Framework for Simulation and Training of Biophysically Detailed NetworksJunfeng Lu, Zijie Yu, Shaoyang Cui, Gan He et al.ICML 2026
