What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
Keyon Vafa, Peter G. Chang, Ashesh Rambachan, Sendhil Mullainathan
摘要
Foundation models are premised on the idea that sequence prediction can uncover deeper domain understanding, much like how Kepler's predictions of planetary motion later led to the discovery of Newtonian mechanics. However, evaluating whether these models truly capture deeper structure remains a challenge. We develop a technique for evaluating foundation models that examines how they adapt to synthetic datasets generated from some postulated world model. Our technique measures whether the foundation model's inductive bias aligns with the world model, and so we refer to it as an inductive bias probe. Across multiple domains, we find that foundation models can excel at their training tasks yet fail to develop inductive biases towards the underlying world model when adapted to new tasks. We particularly find that foundation models trained on orbital trajectories consistently fail to apply Newtonian mechanics when adapted to new physics tasks. Further analysis reveals that these models behave as if they develop task-specific heuristics that fail to generalize.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Language Models Struggle to Use Representations Learned In-ContextMichael A. Lepori, Tal Linzen, Ann Yuan, Katja FilippovaACL 2026 · 被引用 3 次
- Softplus Attention with Re-weighting Boosts Length Extrapolation in Large Language ModelsBo Gao, Michael Spratling, Letizia GionfridaICML 2026 · 被引用 1 次
- Learning Task-Sufficient World Models by Synergizing Agentic Exploration and Structured ModelingFan Feng, Yujia Zheng, Minghao Fu, Yongqiang Chen 等ICML 2026
- Convergent World Representations and Divergent TasksCore Francisco ParkICML 2026
- Learning to Extrapolate to New Tasks: A Relational Approach to Task ExtrapolationAdam Ousherovitch, Yixin WangICML 2026
它引用的顶会 Paper17
- Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space DualityTri Dao, Albert GuICML 2024 · 被引用 1,407 次
- Sparse Autoencoders Find Highly Interpretable Features in Language ModelsRobert Huben, Hoagy Cunningham, Logan Riggs Smith, Aidan Ewart 等ICLR 2024 · 被引用 1,072 次
- Discovering Symbolic Models from Deep Learning with Inductive BiasesMiles D. Cranmer, Alvaro Sanchez-Gonzalez, Peter W. Battaglia, Rui Xu 等NeurIPS 2020 · 被引用 736 次
- A Meta-Transfer Objective for Learning to Disentangle Causal MechanismsYoshua Bengio, Tristan Deleu, Nasim Rahaman, Nan Rosemary Ke 等ICLR 2020 · 被引用 371 次
- Language Models Represent Space and TimeWes Gurnee, Max TegmarkICLR 2024 · 被引用 303 次
相关 Paper
- Zero-shot forecasting of chaotic systemsYuanzhao Zhang, William GilpinICLR 2025
- Understanding the Implicit Biases of Design Choices for Time Series Foundation ModelsAnnan Yu, Danielle C. Maddix, Boran Han, Xiyuan Zhang 等ICLR 2026 · 被引用 11 次
- From Kepler to Newton: Inductive Biases Guide Learned World Models in TransformersZiming Liu, Surya Ganguli, Andreas ToliasICML 2026
- In-Context Fine-Tuning for Time-Series Foundation ModelsMatthew Faw, Rajat Sen, Yichen Zhou, Abhimanyu DasICML 2025
- The Perception–Physics Paradox: Probing Scientific Alignment with TC-BenchDingling Yao, Andrea Polesello, Adeel Pervez, Caroline Muller 等ICML 2026
