PLaID++: A Preference Aligned Language Model for Targeted Inorganic Materials Design
Andy Xu, Rohan Desai, Larry Wang, Ethan Ritz, Gabriel Hope
摘要
Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a promising approach to improve correctness in LLMs, however, in many scientific problems, the objective is not necessarily to produce the correct answer, but instead to produce a diverse array of candidates which satisfy a set of constraints. We study this challenge in the context of materials generation. To this end, we introduce PLaID++, an LLM post-trained for stable and property-guided crystal generation. We find that applying naive preference optimization to a coordinate-based crystal representation leads to mode collapse. Hence, we introduce a compact, symmetry-informed Wyckoff text representation which improves computational efficiency and encourages generalization from physical priors. By encoding symmetry constraints directly into text and guiding model outputs towards desirable chemical space, PLaID++ generates structures that are thermodynamically stable, unique, and novel at a 50% greater rate than prior methods. We further demonstrate that unified training across conditional and unconditional tasks are mutually beneficial in data-sparse regimes. Our work demonstrates the potential of adapting post-training techniques from natural language processing to materials design, paving the way for targeted and efficient discovery of novel materials.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Crystal Diffusion Variational Autoencoder for Periodic Material GenerationTian Xie, Xiang Fu, Octavian-Eugen Ganea, Regina Barzilay 等ICLR 2022 · 被引用 394 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
- Crystal Structure Prediction by Joint Equivariant DiffusionRui Jiao, Wenbing Huang, Peijia Lin, Jiaqi Han 等NeurIPS 2023 · 被引用 245 次
相关 Paper
- Fine-Tuned Language Models Generate Stable Inorganic Materials as TextNate Gruver, Anuroop Sriram, Andrea Madotto, Andrew Gordon Wilson 等ICLR 2024 · 被引用 120 次
- CrystalICL: Enabling In-Context Learning for Crystal GenerationRuobing Wang, Qiaoyu Tan, Yili Wang, Ying Wang 等EMNLP 2025 · 被引用 3 次
- Open Materials Generation with Inference-Time Reinforcement LearningPhilipp Höllmer, Stefano MartinianiICML 2026 · 被引用 3 次
- Wyckoff Transformer: Generation of Symmetric CrystalsNikita Kazeev, Wei Nong, Ignat Romanov, Ruiming Zhu 等ICML 2025
- LLM Meets Diffusion: A Hybrid Framework for Crystal Material GenerationSubhojyoti Khastagir, Kishalay Das, Pawan Goyal, Seung-Cheol Lee 等NeurIPS 2025 · 被引用 14 次
