Towards Fast, Specialized Machine Learning Force Fields: Distilling Foundation Models via Energy Hessians
Ishan Amin, Sanjeev Raja, Aditi S. Krishnapriyan
摘要
The foundation model (FM) paradigm is transforming Machine Learning Force Fields (MLFFs), leveraging general-purpose representations and scalable training to perform a variety of computational chemistry tasks. Although MLFF FMs have begun to close the accuracy gap relative to first-principles methods, there is still a strong need for faster inference speed. Additionally, while research is increasingly focused on general-purpose models which transfer across chemical space, practitioners typically only study a small subset of systems at a given time. This underscores the need for fast, specialized MLFFs relevant to specific downstream applications, which preserve test-time physical soundness while maintaining train-time scalability. In this work, we introduce a method for transferring general-purpose representations from MLFF foundation models to smaller, faster MLFFs specialized to specific regions of chemical space. We formulate our approach as a knowledge distillation procedure, where the smaller "student" MLFF is trained to match the Hessians of the energy predictions of the "teacher" foundation model. Our specialized MLFFs can be up to 20 faster than the original foundation model, while retaining, and in some cases exceeding, its performance and that of undistilled models. We also show that distilling from a teacher model with a direct force parameterization into a student model trained with conservative forces (i.e., computed as derivatives of the potential energy) successfully leverages the representations from the large-scale teacher for improved accuracy, while maintaining energy conservation during test-time molecular dynamics simulations. More broadly, our work suggests a new paradigm for MLFF development, in which foundation models are released along with smaller, specialized simulation "engines" for common chemical subsets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 被引用 10 次
- Learning Smooth and Expressive Interatomic Potentials for Physical Property PredictionXiang Fu, Brandon M. Wood, Luis Barroso-Luque, Daniel S. Levine 等ICML 2025
- Speculative Sampling For Faster Molecular DynamicsArthur Kosmala, Stephan Günnemann, Meng Gao, Brandon WoodICML 2026
- PFT: Phonon Fine-tuning for Machine Learned Interatomic PotentialsTeddy Koker, Abhijeet Gangan, Mit Kotak, Jaime Marian 等ICML 2026
- Elign: Equivariant Diffusion Model Alignment from Foundational Machine Learned Force FieldsYunyang Li, Lin Huang, Luojia Xia, Wenhe Zhang 等ICML 2026
它引用的顶会 Paper11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- GemNet: Universal Directional Graph Neural Networks for MoleculesJohannes Gasteiger, Florian Becker, Stephan GünnemannNeurIPS 2021 · 被引用 665 次
- Understanding Knowledge Distillation in Non-autoregressive Machine TranslationChunting Zhou, Jiatao Gu, Graham NeubigICLR 2020 · 被引用 235 次
- Smooth, exact rotational symmetrization for deep learning on point cloudsSergey Pozdnyakov, Michele CeriottiNeurIPS 2023 · 被引用 68 次
相关 Paper
- Machine Learning Force Fields with Data Cost Aware TrainingAlexander Bukharin, Tianyi Liu, Shengjie Wang, Simiao Zuo 等ICML 2023 · 被引用 1 次
- Accelerating Molecular Graph Neural Networks via Knowledge DistillationFilip Ekström Kelvinius, Dimitar Georgiev, Artur P. Toshev, Johannes GasteigerNeurIPS 2023 · 被引用 22 次
- The dark side of the forces: assessing non-conservative force models for atomistic machine learningFilippo Bigi, Marcel F. Langer, Michele CeriottiICML 2025
- Foundry: Distilling 3D Foundation Models for the EdgeGuillaume Letellier, Siddharth Srivastava, Frédéric Jurie, Gaurav SharmaCVPR 2026 · 被引用 1 次
- Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer BehaviorShashank Subramanian, Peter Harrington, Kurt Keutzer, Wahid Bhimji 等NeurIPS 2023 · 被引用 173 次
