A Self-Attention Ansatz for Ab-initio Quantum Chemistry
Ingrid von Glehn, James S. Spencer, David Pfau
摘要
DeepMind, London N1C 4DL, United Kingdom * We present a novel neural network architecture using self-attention, the Wavefunction Transformer (Psiformer), which can be used as an approximation (or Ansatz) for solving the many-electron Schrödinger equation, the fundamental equation for quantum chemistry and material science. This equation can be solved from first principles, requiring no external training data. In recent years, deep neural networks like the FermiNet and PauliNet have been used to significantly improve the accuracy of these first-principle calculations, but they lack an attention-like mechanism for gating interactions between electrons. Here we show that the Psiformer can be used as a drop-in replacement for these other neural networks, often dramatically improving the accuracy of the calculations. On larger molecules especially, the ground state energy can be improved by dozens of kcal/mol, a qualitative leap over previous methods. This demonstrates that self-attention networks can learn complex quantum mechanical correlations between electrons, and are a promising route to reaching unprecedented accuracy in chemical calculations on larger systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Generalizing Neural Wave FunctionsNicholas Gao, Stephan GünnemannICML 2023 · 被引用 38 次
- Wasserstein Quantum Monte Carlo: A Novel Approach for Solving the Quantum Many-Body Schrödinger EquationKirill Neklyudov, Jannes Nys, Luca A. Thiede, Juan Carrasquilla 等NeurIPS 2023 · 被引用 28 次
- Neural Pfaffians: Solving Many Many-Electron Schrödinger EquationsNicholas Gao, Stephan GünnemannNeurIPS 2024 · 被引用 18 次
- Variational Monte Carlo on a Budget - Fine-tuning pre-trained Neural WavefunctionsMichael Scherbela, Leon Gerard, Philipp GrohsNeurIPS 2023 · 被引用 13 次
- Quadratic Quantum Variational Monte CarloBaiyu Su, Qiang LiuNeurIPS 2024 · 被引用 3 次
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals 等ICML 2021 · 被引用 1,399 次
- Large Batch Optimization for Deep Learning: Training BERT in 76 minutesYang You, Jing Li, Sashank J. Reddi, Jonathan Hseu 等ICLR 2020 · 被引用 1,170 次
相关 Paper
- WF-Bench: A Benchmark for Neural-Network WaveFunction Expressivity and Scaling LawsLixing Zhang, Guijing Duan, Di LuoICML 2026
- Enhancing the Scalability and Applicability of Kohn-Sham Hamiltonians for Molecular SystemsYunyang Li, Zaishuo Xia, Lin Huang, Xinran Wei 等ICLR 2025
- Quantum Transformer for Molecular Learning: Multi-Configuration Ground-State Energy PredictionYuichi Kamata, Quoc Hoan Tran, Yasuhiro Endo, Hirotaka OshimaAAAI 2026 · 被引用 4 次
- NNQS-Transformer: an Efficient and Scalable Neural Network Quantum States Approach for Ab initio Quantum ChemistryYangjun Wu, Chu Guo, Yi Fan, Pengyu Zhou 等SC 2023 · 被引用 33 次
- Ab-Initio Potential Energy Surfaces by Pairing GNNs with Neural Wave FunctionsNicholas Gao, Stephan GünnemannICLR 2022 · 被引用 52 次
