On Transferring Transferability: Towards a Theory for Size Generalization
Eitan Levin, Yuxin Ma, Mateo Díaz, Soledad Villar
Abstract
Many modern learning tasks require models that can take inputs of varying sizes. Consequently, dimension-independent architectures have been proposed for domains where the inputs are graphs, sets, and point clouds. Recent work on graph neural networks has explored whether a model trained on low-dimensional data can transfer its performance to higher-dimensional inputs. We extend this body of work by introducing a general framework for transferability across dimensions. We show that transferability corresponds precisely to continuity in a limit space formed by identifying small problem instances with equivalent large ones. This identification is driven by the data and the learning task. We instantiate our framework on existing architectures, and implement the necessary changes to ensure their transferability. Finally, we provide design principles for designing new transferable models. Numerical experiments support our findings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e544a188-b3f8-4af0-a159-9c23ab69da02Cited by top-tier papers7
- On the Expressive Power of GNNs to Solve Linear SDPsChendi Qian, Christopher MorrisICML 2026 · 1 citation
- pscaling small models: Principled warm starts and hyperparameter transferYuxin Ma, Nan Chen, Mateo D Diaz, Soufiane Hayou et al.ICML 2026 · 1 citation
- Size Transferability of Graph Convolutional Networks across Sparsity: A Generalized Graphon PerspectiveQinji Shu, Hang Sheng, Feng Ji, Hui Feng et al.ICML 2026
- Which Algorithms Can Graph Neural Networks Learn?Solveig Wittig, Antonis Vasileiou, Robert R. Nerem, Timo Stoll et al.ICML 2026
- Learning to Approximate Uniform Facility Location via Graph Neural NetworksChendi Qian, Christopher Morris, Stefanie Jegelka, Christian SohlerICML 2026
Builds on23
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- Generalization and Representational Limits of Graph Neural NetworksVikas K. Garg, Stefanie Jegelka, Tommi S. JaakkolaICML 2020 · 363 citations
- The Lipschitz Constant of Self-AttentionHyunjik Kim, George Papamakarios, Andriy MnihICML 2021 · 208 citations
- Graphon Neural Networks and the Transferability of Graph Neural NetworksLuana Ruiz, Luiz F. O. Chamon, Alejandro RibeiroNeurIPS 2020 · 188 citations
- Scalars are universal: Equivariant machine learning, structured like classical physicsSoledad Villar, David W. Hogg, Kate Storey-Fisher, Weichi Yao et al.NeurIPS 2021 · 185 citations
Related papers
- Any-dimensional invariant universalityShengtai Yao, Eitan Levin, Mateo D DiazICML 2026
- Graph Neural Networks Are Not Continuous Across Graph ResolutionsChristian Koke, Yuesong Shen, Abhishek Saroha, Marvin Eisenberger et al.ICML 2026 · 1 citation
- Graph-Structured Gaussian Processes for Transferable Graph LearningJun Wu, Lisa Ainsworth, Andrew Leakey, Haixun Wang et al.NeurIPS 2023 · 2 citations
- GraphBridge: Towards Arbitrary Transfer Learning in GNNsLi Ju, Xingyi Yang, Qi Li, Xinchao WangICLR 2025
- Bridging Input Feature Spaces Towards Graph Foundation ModelsMoshe Eliasof, Krishna Sri Ipsit Mantri, Beatrice Bevilacqua, Bruno Ribeiro et al.ICLR 2026 · 4 citations
