Mapping Networks
Lord Sen, Shyamapada Mukherjee
摘要
The escalating parameter counts in modern deep learning models pose a fundamental challenge to efficient training and resolution of overfitting. We address this by introducing the Mapping Networks which replace the high dimensional weight space by a compact, trainable latent vector based on the hypothesis that the trained parameters of large networks reside on smooth, low-dimensional manifolds. Henceforth, the Mapping Theorem enforced by a dedicated Mapping Loss, shows the existence of a mapping from this latent space to the target weight space both theoretically and in practice. Mapping Networks significantly reduce overfitting and achieve comparable to better performance than target network across complex vision and sequence tasks, including Image Classification, Deepfake Detection etc, with 99.5%, i.e., around 500× reduction in trainable parameters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess 等ICCV 2019 · 被引用 2,966 次
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 被引用 750 次
- Scale-Space Hypernetworks for Efficient Biomedical Image AnalysisJose Javier Gonzalez Ortiz, John V. Guttag, Adrian V. DalcaNeurIPS 2023 · 被引用 1 次
- Low-Rank Compression of Neural Nets: Learning the Rank of Each LayerYerlan Idelbayev, Miguel Á. Carreira-PerpiñánCVPR 2020
- Celeb-DF: A Large-Scale Challenging Dataset for DeepFake ForensicsYuezun Li, Xin Yang, Pu Sun, Honggang Qi 等CVPR 2020
相关 Paper
- Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual NetworksZixuan Zhang, Kaiqi Zhang, Minshuo Chen, Yuma Takeda 等NeurIPS 2024 · 被引用 6 次
- Approximating Latent Manifolds in Neural Networks via Vanishing IdealsNico Pelleriti, Max Zimmer, Elias Samuel Wirth, Sebastian PokuttaICML 2025
- Lookup multivariate Kolmogorov-Arnold NetworksSergey Pozdnyakov, Philippe SchwallerICLR 2026 · 被引用 1 次
- MCNC: Manifold-Constrained Reparameterization for Neural CompressionChayne Thrash, Reed Andreas, Ali Abbasi, Parsa Nooralinejad 等ICLR 2025
- Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural NetworksRunpei Dong, Zhanhong Tan, Mengdi Wu, Linfeng Zhang 等ICML 2022 · 被引用 15 次
