GFFMERGE: Efficient Merging of Graph Neural Force Fields and Beyond
Parth Verma, Parv P Singh, Vipul Garg, Ishita Thakre, N M Anoop Krishnan, Sayan Ranu
摘要
Graph Neural Networks (GNNs) have revolutionized Neural Force Fields for atomistic simulations, achieving near-quantum accuracy at reduced cost, yet adapting these models to new chemical systems requires expensive retraining of foundation models. Inspired by model merging in vision and language processing, we introduce GFFMERGE, the first principled framework for closed-form model merging in GNN force fields, and GNNs in general. We exploit the linear structure of message-passing layers and formulate merging as a convex embedding-alignment problem with an analytical solution. Through the first systematic benchmarking of model merging for GNN force fields, we show that existing methods designed for vision and language catastrophically fail on force field regression, while GFFMERGE recovers performance approaching gold standard joint training. Across molecular (MD17, MD22), solidstate (LiPS20), and large-scale graph benchmarks, GFFMERGE and GNNMERGE (its generic GNN counterpart) achieve 5-27× speedups while enabling modular composition of specialized models. Remarkably, our closed-form solution alone outperforms all baseline methods before finetuning and provides superior initialization for faster, data-efficient convergence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- TIES-Merging: Resolving Interference When Merging ModelsPrateek Yadav, Derek Tam, Leshem Choshen, Colin A. Raffel 等NeurIPS 2023 · 被引用 999 次
- Merging Models with Fisher-Weighted AveragingMichael Matena, Colin RaffelNeurIPS 2022 · 被引用 741 次
- NodeFormer: A Scalable Graph Structure Learning Transformer for Node ClassificationQitian Wu, Wentao Zhao, Zenan Li, David P. Wipf 等NeurIPS 2022 · 被引用 472 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
相关 Paper
- Generalizing Neural Wave FunctionsNicholas Gao, Stephan GünnemannICML 2023 · 被引用 38 次
- On the Scalability of GNNs for Molecular GraphsMaciej Sypetkowski, Frederik Wenkel, Farimah Poursafaei, Nia Dickson 等NeurIPS 2024 · 被引用 58 次
- Rapid and Precise Topological Comparison with Merge Tree Neural NetworksYu Qin, Brittany Terese Fasy, Carola Wenk, Brian SummaIEEE VIS 2024 · 被引用 5 次
- Long-Short-Range Message-Passing: A Physics-Informed Framework to Capture Non-Local Interaction for Scalable Molecular Dynamics SimulationYunyang Li, Yusong Wang, Lin Huang, Han Yang 等ICLR 2024 · 被引用 31 次
- Expressivity-Preserving GNN SimulationFabian Jogl, Maximilian Thiessen, Thomas GärtnerNeurIPS 2023 · 被引用 11 次
