Learning to Extrapolate: A Transductive Approach
Aviv Netanyahu, Abhishek Gupta, Max Simchowitz, Kaiqing Zhang, Pulkit Agrawal
摘要
Machine learning systems, especially with overparameterized deep neural networks, can generalize to novel test instances drawn from the same distribution as the training data. However, they fare poorly when evaluated on out-of-support test points. In this work, we tackle the problem of developing machine learning systems that retain the power of overparameterized function approximators while enabling extrapolation to out-of-support test points when possible. This is accomplished by noting that under certain conditions, a "transductive" reparameterization can convert an out-of-support extrapolation problem into a problem of within-support combinatorial generalization. We propose a simple strategy based on bilinear embeddings to enable this type of combinatorial generalization, thereby addressing the out-of-support extrapolation problem under certain conditions. We instantiate a simple, practical algorithm applicable to various supervised learning and imitation learning tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Compositional Generalization from First PrinciplesThaddäus Wiedemer, Prasanna Mayilvahanan, Matthias Bethge, Wieland BrendelNeurIPS 2023 · 被引用 78 次
- Provable Compositional Generalization for Object-Centric LearningThaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos 等ICLR 2024 · 被引用 40 次
- Accurate and Scalable Estimation of Epistemic Uncertainty for Graph Neural NetworksPuja Trivedi, Mark Heimann, Rushil Anirudh, Danai Koutra 等ICLR 2024 · 被引用 8 次
- Towards Understanding Extrapolation: a Causal LensLingjing Kong, Guangyi Chen, Petar Stojanov, Haoxuan Li 等NeurIPS 2024 · 被引用 7 次
- PAGER: Accurate Failure Characterization in Deep Regression ModelsJayaraman J. Thiagarajan, Vivek Sivaraman Narayanaswamy, Puja Trivedi, Rushil AnirudhICML 2024 · 被引用 5 次
它引用的顶会 Paper19
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 被引用 2,878 次
- Solving Quantitative Reasoning Problems with Language ModelsAitor Lewkowycz, Anders Andreassen, David Dohan, Ethan Dyer 等NeurIPS 2022 · 被引用 2,039 次
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
相关 Paper
- VIKING: Deep variational inference with stochastic projectionsSamuel Matthiesen, Hrittik Roy, Nicholas Krämer, Yevgen Zainchkovskyy 等NeurIPS 2025 · 被引用 3 次
- Graph Structure Extrapolation for Out-of-Distribution GeneralizationXiner Li, Shurui Gui, Youzhi Luo, Shuiwang JiICML 2024 · 被引用 12 次
- Nonlinearity Encoding for Extrapolation of Neural NetworksGyoung S. Na, Chanyoung ParkKDD 2022 · 被引用 4 次
- Learning Representations that Support ExtrapolationTaylor W. Webb, Zachary Dulberg, Steven Frankland, Alexander A. Petrov 等ICML 2020 · 被引用 60 次
- Learning to Extrapolate to New Tasks: A Relational Approach to Task ExtrapolationAdam Ousherovitch, Yixin WangICML 2026
