Fine-Tuned Language Models Generate Stable Inorganic Materials as Text
Nate Gruver, Anuroop Sriram, Andrea Madotto, Andrew Gordon Wilson, C. Lawrence Zitnick, Zachary W. Ulissi
摘要
We propose fine-tuning large language models for generation of stable materials. While unorthodox, fine-tuning large language models on text-encoded atomistic data is simple to implement yet reliable, with around 90% of sampled structures obeying physical constraints on atom positions and charges. Using energy above hull calculations from both learned ML potentials and gold-standard DFT calculations, we show that our strongest model (fine-tuned LLaMA-2 70B) can generate materials predicted to be metastable at about twice the rate (49% vs 28%) of CD-VAE, a competing diffusion model. Because of text prompting's inherent flexibility, our models can simultaneously be used for unconditional generation of stable material, infilling of partial structures and text-conditional generation. Finally, we show that language models' ability to capture key symmetries of crystal structures improves with model scale, suggesting that the biases of pretrained LLMs are surprisingly well-suited for atomistic data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- FlowMM: Generating Materials with Riemannian Flow MatchingBenjamin Kurt Miller, Ricky T. Q. Chen, Anuroop Sriram, Brandon M. WoodICML 2024 · 被引用 101 次
- FlowLLM: Flow Matching for Material Generation with Large Language Models as Base DistributionsAnuroop Sriram, Benjamin Kurt Miller, Ricky T. Q. Chen, Brandon M. WoodNeurIPS 2024 · 被引用 78 次
- A Comprehensive Survey of Scientific Large Language Models and Their Applications in Scientific DiscoveryYu Zhang, Xiusi Chen, Bowen Jin, Sheng Wang 等EMNLP 2024 · 被引用 28 次
- Generative Hierarchical Materials SearchSherry Yang, Simon L. Batzner, Ruiqi Gao, Muratahan Aykol 等NeurIPS 2024 · 被引用 23 次
- Invariant Tokenization of Crystalline Materials for Language Model Enabled GenerationKeqiang Yan, Xiner Li, Hongyi Ling, Kenna Ashen 等NeurIPS 2024 · 被引用 23 次
它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung 等ICLR 2020 · 被引用 1,166 次
- Large Language Models Are Zero-Shot Time Series ForecastersNate Gruver, Marc Finzi, Shikai Qiu, Andrew Gordon WilsonNeurIPS 2023 · 被引用 898 次
相关 Paper
- LLM Meets Diffusion: A Hybrid Framework for Crystal Material GenerationSubhojyoti Khastagir, Kishalay Das, Pawan Goyal, Seung-Cheol Lee 等NeurIPS 2025 · 被引用 14 次
- CrystalICL: Enabling In-Context Learning for Crystal GenerationRuobing Wang, Qiaoyu Tan, Yili Wang, Ying Wang 等EMNLP 2025 · 被引用 3 次
- PLaID++: A Preference Aligned Language Model for Targeted Inorganic Materials DesignAndy Xu, Rohan Desai, Larry Wang, Ethan Ritz 等ICML 2026 · 被引用 11 次
- Scalable Diffusion for Materials GenerationSherry Yang, KwangHwan Cho, Amil Merchant, Pieter Abbeel 等ICLR 2024 · 被引用 77 次
- Free and Fair Hardware: A Pathway to Copyright Infringement-Free Verilog Generation using LLMsSam Bush, Matthew DeLorenzo, Phat Tieu, Jeyavijayan RajendranDAC 2025 · 被引用 5 次
