UMA: A Family of Universal Models for Atoms
Brandon M. Wood, Misko Dzamba, Xiang Fu, Meng Gao, Muhammed Shuaibi, Luis Barroso-Luque, Kareem Abdelmaqsoud, Vahe Gharakhanyan, John R. Kitchin, Daniel S. Levine, Kyle Michel, Anuroop Sriram
Abstract
The ability to quickly and accurately compute properties from atomic simulations is critical for advancing a large number of applications in chemistry and materials science including drug discovery, energy storage, and semiconductor manufacturing. To address this need, Meta FAIR presents a family of Universal Models for Atoms (UMA), designed to push the frontier of speed, accuracy, and generalization. UMA models are trained on half a billion unique 3D atomic structures (the largest training runs to date) by compiling data across multiple chemical domains, e.g. molecules, materials, and catalysts. We develop empirical scaling laws to help understand how to increase model capacity alongside dataset size to achieve the best accuracy. The UMA small and medium models utilize a novel architectural design we refer to as mixture of linear experts that enables increasing model capacity without sacrificing speed. For example, UMA-medium has 1.4B parameters but only 50M active parameters per atomic structure. We evaluate UMA models on a diverse set of applications across multiple domains and find that, remarkably, a single model without any fine-tuning can perform similarly or better than specialized models. We are releasing the UMA code, weights, and associated data to accelerate computational workflows and enable the community to continue to build increasingly capable AI models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 057cd86b-47b1-42d2-a680-0effd24da296Cited by top-tier papers13
- Mol-LLaMA: Towards General Understanding of Molecules in Large Molecular Language ModelDongki Kim, Wonbin Lee, Sung Ju HwangNeurIPS 2025 · 24 citations
- Enhancing Diffusion-Based Sampling with Molecular Collective VariablesJuno Nam, Bálint Máté, Artur P. Toshev, Manasa Kaniselvan et al.ICLR 2026 · 15 citations
- Flexible MOF Generation with Torsion-Aware Flow MatchingNayoung Kim, Seongsu Kim, Sungsoo AhnNeurIPS 2025 · 13 citations
- OXtal: An All-Atom Diffusion Model for Organic Crystal Structure PredictionEmily Jin, Andrei Cristian Nica, Mikhail Galkin, Jarrid Rector-Brooks et al.ICLR 2026 · 9 citations
- Learning from the Electronic Structure of Molecules across the Periodic TableManasa Kaniselvan, Benjamin Kurt Miller, Meng Gao, Juno Nam et al.ICLR 2026 · 8 citations
Builds on9
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- Scaling Vision with Sparse Mixture of ExpertsCarlos Riquelme, Joan Puigcerver, Basil Mustafa, Maxim Neumann et al.NeurIPS 2021 · 1,213 citations
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 311 citations
- DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language ModelsDamai Dai, Chengqi Deng, Chenggang Zhao, R. X. Xu et al.ACL 2024 · 171 citations
- Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNsSaro Passaro, C. Lawrence ZitnickICML 2023 · 157 citations
Related papers
- Scaling Laws of Graph Neural Networks for Atomistic Materials ModelingChaojian Li, Zhifan Ye, Massimiliano Lupo Pasini, Jong Youl Choi et al.DAC 2025 · 2 citations
- Scaling Laws and Symmetry, Evidence from Neural Force FieldsNhat Khang Ngo, Siamak RavanbakhshICLR 2026 · 7 citations
- All-atom Diffusion Transformers: Unified generative modelling of molecules and materialsChaitanya K. Joshi, Xiang Fu, Yi-Lun Liao, Vahe Gharakhanyan et al.ICML 2025
- Exploring Molecular Pretraining Model at ScaleXiaohong Ji, Zhen Wang, Zhifeng Gao, Hang Zheng et al.NeurIPS 2024 · 24 citations
- ATOM: A Pretrained Neural Operator for Multitask Molecular DynamicsLuke Thompson, Davy Guan, Slade Matthews, Dai Shi et al.ICLR 2026 · 1 citation
