From Kepler to Newton: Inductive Biases Guide Learned World Models in Transformers
Ziming Liu, Surya Ganguli, Andreas Tolias
摘要
Can general-purpose AI architectures go beyond prediction to discover the physical laws governing the universe? True intelligence relies on "world models"-causal abstractions that allow an agent to not only predict future states but understand the underlying governing dynamics. While previous "AI Physicist" approaches have successfully recovered such laws, they typically rely on strong, domain-specific priors that effectively "bake in" the physics. Conversely, Vafa et al. ( 2025 ) recently showed that generic Transformers fail to acquire these world models, achieving high predictive accuracy without capturing the underlying physical laws. We bridge this gap by systematically introducing three minimal inductive biases. We show that ensuring spatial smoothness (by formulating prediction as continuous regression) and stability (by training with noisy contexts to mitigate error accumulation) enables generic Transformers to surpass prior failures and learn a coherent Keplerian world model, successfully fitting ellipses to planetary trajectories. However, true physical insight requires a third bias: temporal locality. By restricting the attention window to the immediate past-imposing the simple assumption that future states depend only on the local state rather than a complex history-we force the model to abandon curve-fitting and discover Newtonian force representations. Our results demonstrate that simple architectural choices determine whether an AI becomes a curve-fitter or a physicist, marking a critical step toward automated scientific discovery.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech 等NeurIPS 2022 · 被引用 6,707 次
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 被引用 1,852 次
相关 Paper
- What Has a Foundation Model Found? Using Inductive Bias to Probe for World ModelsKeyon Vafa, Peter G. Chang, Ashesh Rambachan, Sendhil MullainathanICML 2025
- Deconstructing the Inductive Biases of Hamiltonian Neural NetworksNate Gruver, Marc Anton Finzi, Samuel Don Stanton, Andrew Gordon WilsonICLR 2022 · 被引用 50 次
- Physics-Integrated Variational Autoencoders for Robust and Interpretable Generative ModelingNaoya Takeishi, Alexandros KalousisNeurIPS 2021 · 被引用 88 次
- Interpreting Physics in Video World ModelsSonia Joseph, Quentin Garrido, Randall Balestriero, Matthew Kowal 等ICML 2026
- Physics-as-Inverse-Graphics: Unsupervised Physical Parameter Estimation from VideoMiguel Jaques, Michael Burke, Timothy M. HospedalesICLR 2020 · 被引用 58 次
