Functional Indirection Neural Estimator for Better Out-of-distribution Generalization
Kha Pham, Hung Le, Man Ngo, Truyen Tran
Abstract
The capacity to achieve out-of-distribution (OOD) generalization is a hallmark of human intelligence and yet remains out of reach for machines. This remarkable capability has been attributed to our abilities to make conceptual abstraction and analogy, and to a mechanism known as indirection, which binds two representations and uses one representation to refer to the other. Inspired by these mechanisms, we hypothesize that OOD generalization may be achieved by performing analogy-making and indirection in the functional space instead of the data space as in current methods. To realize this, we design FINE (Functional Indirection Neural Estimator), a neural framework that learns to compose functions that map data input to output on-the-fly. FINE consists of a backbone network and a trainable semantic memory of basis weight matrices. Upon seeing a new input-output data pair, FINE dynamically constructs the backbone weights by mixing the basis weights. The mixing coefficients are indirectly computed through querying a separate corresponding semantic memory using the data pair. We demonstrate empirically that FINE can strongly improve out-of-distribution generalization on IQ tasks that involve geometric transformations. In particular, we train FINE and competing models on IQ tasks using images from the MNIST, Omniglot and CIFAR100 datasets and test on tasks with unseen image classes from one or different datasets and unseen transformation rules. FINE not only achieves the best performance on all tasks but also is able to adapt to small-scale data scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e8a73621-c873-48d2-bd64-285aa68ac440Builds on6
- An Explicitly Relational Neural Network ArchitectureMurray Shanahan, Kyriacos Nikiforou, Antonia Creswell, Christos Kaplanis et al.ICML 2020 · 72 citations
- Emergent Symbols through Binding in External MemoryTaylor Whittington Webb, Ishan Sinha, Jonathan D. CohenICLR 2021 · 68 citations
- Learning Representations that Support ExtrapolationTaylor W. Webb, Zachary Dulberg, Steven Frankland, Alexander A. Petrov et al.ICML 2020 · 60 citations
- The Devil is in the Detail: Simple Tricks Improve Systematic Generalization of TransformersRóbert Csordás, Kazuki Irie, Jürgen SchmidhuberEMNLP 2021 · 55 citations
- Neural Stored-program MemoryHung Le, Truyen Tran, Svetha VenkateshICLR 2020 · 38 citations
Related papers
- Improving Out-of-distribution Generalization with Indirection RepresentationsKha Pham, Hung Le, Man Ngo, Truyen TranICLR 2023
- Drawing out of Distribution with Neuro-Symbolic Generative ModelsYichao Liang, Josh Tenenbaum, Tuan Anh Le, N. SiddharthNeurIPS 2022 · 12 citations
- Combining Induction and Transduction for Abstract ReasoningWen-Ding Li, Keya Hu, Carter Larsen, Yuqing Wu et al.ICLR 2025
- Emergent Analogical Reasoning in TransformersGouki Minegishi, Jingyuan Feng, Hiroki Furuta, Takeshi Kojima et al.ICML 2026 · 4 citations
- Generalizing Analogical Inference from Boolean to Continuous DomainsFrancisco Cunha, Yves Lepage, Miguel Couceiro, Zied BouraouiAAAI 2026 · 2 citations
