UMA: A Family of Universal Models for Atoms
Brandon M. Wood, Misko Dzamba, Xiang Fu, Meng Gao, Muhammed Shuaibi, Luis Barroso-Luque, Kareem Abdelmaqsoud, Vahe Gharakhanyan, John R. Kitchin, Daniel S. Levine, Kyle Michel, Anuroop Sriram
摘要
The ability to quickly and accurately compute properties from atomic simulations is critical for advancing a large number of applications in chemistry and materials science including drug discovery, energy storage, and semiconductor manufacturing. To address this need, Meta FAIR presents a family of Universal Models for Atoms (UMA), designed to push the frontier of speed, accuracy, and generalization. UMA models are trained on half a billion unique 3D atomic structures (the largest training runs to date) by compiling data across multiple chemical domains, e.g. molecules, materials, and catalysts. We develop empirical scaling laws to help understand how to increase model capacity alongside dataset size to achieve the best accuracy. The UMA small and medium models utilize a novel architectural design we refer to as mixture of linear experts that enables increasing model capacity without sacrificing speed. For example, UMA-medium has 1.4B parameters but only 50M active parameters per atomic structure. We evaluate UMA models on a diverse set of applications across multiple domains and find that, remarkably, a single model without any fine-tuning can perform similarly or better than specialized models. We are releasing the UMA code, weights, and associated data to accelerate computational workflows and enable the community to continue to build increasingly capable AI models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Mol-LLaMA: Towards General Understanding of Molecules in Large Molecular Language ModelDongki Kim, Wonbin Lee, Sung Ju HwangNeurIPS 2025 · 被引用 24 次
- Enhancing Diffusion-Based Sampling with Molecular Collective VariablesJuno Nam, Bálint Máté, Artur P. Toshev, Manasa Kaniselvan 等ICLR 2026 · 被引用 15 次
- Flexible MOF Generation with Torsion-Aware Flow MatchingNayoung Kim, Seongsu Kim, Sungsoo AhnNeurIPS 2025 · 被引用 13 次
- OXtal: An All-Atom Diffusion Model for Organic Crystal Structure PredictionEmily Jin, Andrei Cristian Nica, Mikhail Galkin, Jarrid Rector-Brooks 等ICLR 2026 · 被引用 9 次
- Learning from the Electronic Structure of Molecules across the Periodic TableManasa Kaniselvan, Benjamin Kurt Miller, Meng Gao, Juno Nam 等ICLR 2026 · 被引用 8 次
它引用的顶会 Paper9
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- Scaling Vision with Sparse Mixture of ExpertsCarlos Riquelme, Joan Puigcerver, Basil Mustafa, Maxim Neumann 等NeurIPS 2021 · 被引用 1,213 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
- DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language ModelsDamai Dai, Chengqi Deng, Chenggang Zhao, R. X. Xu 等ACL 2024 · 被引用 171 次
- Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNsSaro Passaro, C. Lawrence ZitnickICML 2023 · 被引用 157 次
相关 Paper
- Scaling Laws of Graph Neural Networks for Atomistic Materials ModelingChaojian Li, Zhifan Ye, Massimiliano Lupo Pasini, Jong Youl Choi 等DAC 2025 · 被引用 2 次
- Scaling Laws and Symmetry, Evidence from Neural Force FieldsNhat Khang Ngo, Siamak RavanbakhshICLR 2026 · 被引用 7 次
- All-atom Diffusion Transformers: Unified generative modelling of molecules and materialsChaitanya K. Joshi, Xiang Fu, Yi-Lun Liao, Vahe Gharakhanyan 等ICML 2025
- Exploring Molecular Pretraining Model at ScaleXiaohong Ji, Zhen Wang, Zhifeng Gao, Hang Zheng 等NeurIPS 2024 · 被引用 24 次
- ATOM: A Pretrained Neural Operator for Multitask Molecular DynamicsLuke Thompson, Davy Guan, Slade Matthews, Dai Shi 等ICLR 2026 · 被引用 1 次
