Controlling the Complexity and Lipschitz Constant improves Polynomial Nets
Zhenyu Zhu, Fabian Latorre, Grigorios Chrysos, Volkan Cevher
摘要
While the class of Polynomial Nets demonstrates comparable performance to neural networks (NN), it currently has neither theoretical generalization characterization nor robustness guarantees. To this end, we derive new complexity bounds for the set of Coupled CP-Decomposition (CCP) and Nested Coupled CP-decomposition (NCP) models of Polynomial Nets in terms of the -operator-norm and the -operator norm. In addition, we derive bounds on the Lipschitz constant for both models to establish a theoretical certificate for their robustness. The theoretical results enable us to propose a principled regularization scheme that we also evaluate experimentally in six datasets and show that it improves the accuracy as well as the robustness of the models to adversarial perturbations. We showcase how this regularization can be combined with adversarial training, resulting in further improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Pay attention to your loss : understanding misconceptions about Lipschitz neural networksLouis Béthune, Thibaut Boissin, Mathieu Serrurier, Franck Mamalet 等NeurIPS 2022 · 被引用 35 次
- Extrapolation and Spectral Bias of Neural Nets with Hadamard Product: a Polynomial Net StudyYongtao Wu, Zhenyu Zhu, Fanghui Liu, Grigorios Chrysos 等NeurIPS 2022 · 被引用 19 次
- Sound and Complete Verification of Polynomial NetworksElías Abad-Rocamora, Mehmet Fatih Sahin, Fanghui Liu, Grigorios Chrysos 等NeurIPS 2022 · 被引用 6 次
- Regularization of polynomial networks for image recognitionGrigorios G. Chrysos, Bohan Wang, Jiankang Deng, Volkan CevherCVPR 2023
- Low-Rank Adaptation in Multilinear Operator Networks for Security-Preserving Incremental LearningHuu Binh Ta, Duc Nguyen, Quyen Tran, Toan Tran 等CVPR 2025
它引用的顶会 Paper4
- Reliable evaluation of adversarial robustness with an ensemble of diverse parameter-free attacksFrancesco Croce, Matthias HeinICML 2020 · 被引用 2,337 次
- Lipschitz constant estimation of Neural Networks via sparse polynomial optimizationFabian Latorre, Paul Rolland, Volkan CevherICLR 2020 · 被引用 154 次
- Multiplicative Interactions and Where to Find ThemSiddhant M. Jayakumar, Wojciech M. Czarnecki, Jacob Menick, Jonathan Schwarz 等ICLR 2020 · 被引用 152 次
- Convolutional Tensor-Train LSTM for Spatio-Temporal LearningJiahao Su, Wonmin Byeon, Jean Kossaifi, Furong Huang 等NeurIPS 2020 · 被引用 146 次
相关 Paper
- Large Norms of CNN Layers Do Not Hurt Adversarial RobustnessYouwei Liang, Dong HuangAAAI 2021 · 被引用 13 次
- On Lipschitz Regularization of Convolutional Layers using Toeplitz Matrix TheoryAlexandre Araujo, Benjamin Négrevergne, Yann Chevaleyre, Jamal AtifAAAI 2021 · 被引用 31 次
- Improved deterministic l2 robustness on CIFAR-10 and CIFAR-100Sahil Singla, Surbhi Singla, Soheil FeiziICLR 2022 · 被引用 77 次
- Adversarial Training is a Form of Data-dependent Operator Norm RegularizationKevin Roth, Yannic Kilcher, Thomas HofmannNeurIPS 2020 · 被引用 61 次
- Adversarial Training and Provable Defenses: Bridging the GapMislav Balunovic, Martin T. VechevICLR 2020 · 被引用 186 次
