Functional Indirection Neural Estimator for Better Out-of-distribution Generalization
Kha Pham, Hung Le, Man Ngo, Truyen Tran
摘要
The capacity to achieve out-of-distribution (OOD) generalization is a hallmark of human intelligence and yet remains out of reach for machines. This remarkable capability has been attributed to our abilities to make conceptual abstraction and analogy, and to a mechanism known as indirection, which binds two representations and uses one representation to refer to the other. Inspired by these mechanisms, we hypothesize that OOD generalization may be achieved by performing analogy-making and indirection in the functional space instead of the data space as in current methods. To realize this, we design FINE (Functional Indirection Neural Estimator), a neural framework that learns to compose functions that map data input to output on-the-fly. FINE consists of a backbone network and a trainable semantic memory of basis weight matrices. Upon seeing a new input-output data pair, FINE dynamically constructs the backbone weights by mixing the basis weights. The mixing coefficients are indirectly computed through querying a separate corresponding semantic memory using the data pair. We demonstrate empirically that FINE can strongly improve out-of-distribution generalization on IQ tasks that involve geometric transformations. In particular, we train FINE and competing models on IQ tasks using images from the MNIST, Omniglot and CIFAR100 datasets and test on tasks with unseen image classes from one or different datasets and unseen transformation rules. FINE not only achieves the best performance on all tasks but also is able to adapt to small-scale data scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- An Explicitly Relational Neural Network ArchitectureMurray Shanahan, Kyriacos Nikiforou, Antonia Creswell, Christos Kaplanis 等ICML 2020 · 被引用 72 次
- Emergent Symbols through Binding in External MemoryTaylor Whittington Webb, Ishan Sinha, Jonathan D. CohenICLR 2021 · 被引用 68 次
- Learning Representations that Support ExtrapolationTaylor W. Webb, Zachary Dulberg, Steven Frankland, Alexander A. Petrov 等ICML 2020 · 被引用 60 次
- The Devil is in the Detail: Simple Tricks Improve Systematic Generalization of TransformersRóbert Csordás, Kazuki Irie, Jürgen SchmidhuberEMNLP 2021 · 被引用 55 次
- Neural Stored-program MemoryHung Le, Truyen Tran, Svetha VenkateshICLR 2020 · 被引用 38 次
相关 Paper
- Improving Out-of-distribution Generalization with Indirection RepresentationsKha Pham, Hung Le, Man Ngo, Truyen TranICLR 2023
- Drawing out of Distribution with Neuro-Symbolic Generative ModelsYichao Liang, Josh Tenenbaum, Tuan Anh Le, N. SiddharthNeurIPS 2022 · 被引用 12 次
- Combining Induction and Transduction for Abstract ReasoningWen-Ding Li, Keya Hu, Carter Larsen, Yuqing Wu 等ICLR 2025
- Emergent Analogical Reasoning in TransformersGouki Minegishi, Jingyuan Feng, Hiroki Furuta, Takeshi Kojima 等ICML 2026 · 被引用 4 次
- Generalizing Analogical Inference from Boolean to Continuous DomainsFrancisco Cunha, Yves Lepage, Miguel Couceiro, Zied BouraouiAAAI 2026 · 被引用 2 次
