Mapping Networks
Lord Sen, Shyamapada Mukherjee
Abstract
The escalating parameter counts in modern deep learning models pose a fundamental challenge to efficient training and resolution of overfitting. We address this by introducing the Mapping Networks which replace the high dimensional weight space by a compact, trainable latent vector based on the hypothesis that the trained parameters of large networks reside on smooth, low-dimensional manifolds. Henceforth, the Mapping Theorem enforced by a dedicated Mapping Loss, shows the existence of a mapping from this latent space to the target weight space both theoretically and in practice. Mapping Networks significantly reduce overfitting and achieve comparable to better performance than target network across complex vision and sequence tasks, including Image Classification, Deepfake Detection etc, with 99.5%, i.e., around 500× reduction in trainable parameters.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
- Scale-Space Hypernetworks for Efficient Biomedical Image AnalysisJose Javier Gonzalez Ortiz, John V. Guttag, Adrian V. DalcaNeurIPS 2023 · 1 citation
- Low-Rank Compression of Neural Nets: Learning the Rank of Each LayerYerlan Idelbayev, Miguel Á. Carreira-PerpiñánCVPR 2020
- Celeb-DF: A Large-Scale Challenging Dataset for DeepFake ForensicsYuezun Li, Xin Yang, Pu Sun, Honggang Qi et al.CVPR 2020
Related papers
- Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual NetworksZixuan Zhang, Kaiqi Zhang, Minshuo Chen, Yuma Takeda et al.NeurIPS 2024 · 6 citations
- Approximating Latent Manifolds in Neural Networks via Vanishing IdealsNico Pelleriti, Max Zimmer, Elias Samuel Wirth, Sebastian PokuttaICML 2025
- Lookup multivariate Kolmogorov-Arnold NetworksSergey Pozdnyakov, Philippe SchwallerICLR 2026 · 1 citation
- MCNC: Manifold-Constrained Reparameterization for Neural CompressionChayne Thrash, Reed Andreas, Ali Abbasi, Parsa Nooralinejad et al.ICLR 2025
- Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural NetworksRunpei Dong, Zhanhong Tan, Mengdi Wu, Linfeng Zhang et al.ICML 2022 · 15 citations
