InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames
Haorui Li, weitao du, Yuqiang Li, Hongyu Guo, Shengchao Liu
摘要
Transformer-based autoregressive models have emerged as a unifying paradigm across modalities such as text and images, but their extension to 3D molecule generation remains underexplored. The gap stems from two fundamental challenges: (1) how to tokenize molecules into a canonical 1D sequence of tokens that is invariant to both SE(3) transformations and atom index permutations, and (2) how to design an architecture capable of modeling hybrid atom-based tokens that couple discrete atom types with continuous 3D coordinates. To address these challenges, we introduce InertialAR. It first performs generation-oriented canonical tokenization by aligning each molecule to a canonical inertial frame and reordering atoms, thereby converting arbitrary 3D structures into a unique, SE(3)- and permutation-invariant sequence of tokens for autoregressive generation. Built upon this canonical tokenization, we propose geometric positional encoding (GeoPE), which endows Transformer attention with 3D geometric awareness. Finally, InertialAR utilizes a hierarchical autoregressive paradigm to decode the next atom, consecutively predicting the atom type and 3D coordinates via Diffusion Loss. Experimentally, InertialAR achieves state-of-the-art performance on 8 of the 10 evaluation metrics for unconditional generation across QM9, GEOM-Drugs, and B3LYP. Moreover, it significantly outperforms baselines in controllable generation for targeted chemical functionality, attaining state-of-the-art results across all 5 metrics. Code is available at github.com/HaoruiLi46/InertialAR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- A Resolution-Agnostic Geometric Transformer for Chromosome Modeling Using Inertial FrameYize Zhou, Haorui Li, Shengchao LiuICLR 2026 · 被引用 2 次
- Rigidity-Aware Geometric Pretraining for Protein Design and Conformational EnsemblesZhanghan Ni, Yanjing Li, Zeju Qiu, Bernhard Schölkopf 等ICLR 2026 · 被引用 2 次
它引用的顶会 Paper23
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 被引用 1,432 次
- Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale PredictionKeyu Tian, Yi Jiang, Zehuan Yuan, Bingyue Peng 等NeurIPS 2024 · 被引用 1,199 次
相关 Paper
- Towards Unified and Lossless Latent Space for 3D Molecular Latent Diffusion ModelingYanchen Luo, Zhiyuan Liu, Yi Zhao, Sihang Li 等NeurIPS 2025 · 被引用 11 次
- Geometric Transformer with Interatomic Positional EncodingYusong Wang, Shaoning Li, Tong Wang, Bin Shao 等NeurIPS 2023 · 被引用 25 次
- Geometry Informed Tokenization of Molecules for Language Model GenerationXiner Li, Limei Wang, Youzhi Luo, Carl Edwards 等ICML 2025
- Tokenizing 3D Molecule Structure with Quantized Spherical CoordinatesKaiyuan Gao, Yusong Wang, Haoxiang Guan, Zun Wang 等KDD 2026 · 被引用 5 次
- Sampling 3D Molecular Conformers with Diffusion TransformersJ. Thorben Frank, Winfried Ripken, Gregor Lied, Klaus-Robert Müller 等NeurIPS 2025 · 被引用 7 次
