Regularization of polynomial networks for image recognition
Grigorios G. Chrysos, Bohan Wang, Jiankang Deng, Volkan Cevher
摘要
Deep Neural Networks (DNNs) have obtained impressive performance across tasks, however they still remain as black boxes, e.g., hard to theoretically analyze. At the same time, Polynomial Networks (PNs) have emerged as an alternative method with a promising performance and improved interpretability but have yet to reach the performance of the powerful DNN baselines. In this work, we aim to close this performance gap. We introduce a class of PNs, which are able to reach the performance of ResNet across a range of six benchmarks. We demonstrate that strong regularization is critical and conduct an extensive study of the exact regularization schemes required to match performance. To further motivate the regularization schemes, we introduce D-PolyNets that achieve a higherdegree of expansion than previously proposed polynomial networks. D-PolyNets are more parameter-efficient while achieving a similar performance as other polynomial networks. We expect that our new models can lead to an understanding of the role of elementwise activation functions (which are no longer required for training PNs). The source code is available at https://github.com/ grigorisg9gr/regularized_polynomials.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Beyond Linear Probes: Dynamic Safety Monitoring for Language ModelsJames Oldfield, Philip Torr, Ioannis Patras, Adel Bibi 等ICLR 2026 · 被引用 16 次
- Deep Tree Tensor NetworksChang NieNeurIPS 2025 · 被引用 4 次
- Polynomial, trigonometric, and tropical activationsIsmail Khalfaoui Hassani, Stefan KesselheimICLR 2026 · 被引用 1 次
- Activation-Free Backbones for Image Recognition: Polynomial Alternatives within MetaFormer-Style Vision ModelsJeffrey Wang, Jonathan Gregory, Grigorios ChrysosICML 2026
- Multilinear Operator NetworksYixin Cheng, Grigorios Chrysos, Markos Georgopoulos, Volkan CevherICLR 2024
它引用的顶会 Paper7
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Scalable Interpretability via PolynomialsAbhimanyu Dubey, Filip Radenovic, Dhruv MahajanNeurIPS 2022 · 被引用 42 次
- Polynomial Neural Fields for Subband Decomposition and ManipulationGuandao Yang, Sagie Benaim, Varun Jampani, Kyle Genova 等NeurIPS 2022 · 被引用 26 次
- The Spectral Bias of Polynomial Neural NetworksMoulik Choraria, Leello Tadesse Dadi, Grigorios Chrysos, Julien Mairal 等ICLR 2022 · 被引用 26 次
- Extrapolation and Spectral Bias of Neural Nets with Hadamard Product: a Polynomial Net StudyYongtao Wu, Zhenyu Zhu, Fanghui Liu, Grigorios Chrysos 等NeurIPS 2022 · 被引用 19 次
相关 Paper
- P-nets: Deep Polynomial Neural NetworksGrigorios G. Chrysos, Stylianos Moschoglou, Giorgos Bouritsas, Yannis Panagakis 等CVPR 2020
- Identifiability of Deep Polynomial Neural NetworksKonstantin Usevich, Ricardo Augusto Borsoi, Clara Dérand, Marianne ClauselNeurIPS 2025 · 被引用 21 次
- Characterizing ResNet's Universal Approximation CapabilityChenghao Liu, Enming Liang, Minghua ChenICML 2024
- ULD-Net: Enabling Ultra-Low-Degree Fully Polynomial Networks for Homomorphically Encrypted InferenceXi Xie, Ran Ran, Jiahui Zhao, Bin Lei 等ICLR 2026
- Controlling the Complexity and Lipschitz Constant improves Polynomial NetsZhenyu Zhu, Fabian Latorre, Grigorios Chrysos, Volkan CevherICLR 2022 · 被引用 12 次
