Adaptive Knowledge Driven Regularization for Deep Neural Networks
Zhaojing Luo, Shaofeng Cai, Can Cui, Beng Chin Ooi, Yang Yang
Abstract
In many real-world applications, the amount of data available for training is often limited, and thus inductive bias and auxiliary knowledge are much needed for regularizing model training. One popular regularization method is to impose prior distribution assumptions on model parameters, and many recent works also attempt to regularize training by integrating external knowledge into specific neurons. However, existing regularization methods fail to take account of the interaction between connected neuron pairs, which is invaluable internal knowledge for adaptive regularization for better representation learning as training progresses. In this paper, we explicitly take into account the interaction between connected neurons, and propose an adaptive internal knowledge driven regularization method, CORR-Reg. The key idea of CORR-Reg is to give a higher significance weight to connections of more correlated neuron pairs. The significance weights adaptively identify more important input neurons for each neuron. Instead of regularizing connection model parameters with a static strength such as weight decay, CORR-Reg imposes weaker regularization strength on more significant connections. As a consequence, neurons attend to more informative input features and thus learn more diversified and discriminative representation. We derive CORR-Reg with the Bayesian inference framework and propose a novel optimization algorithm with the Lagrange multiplier method and Stochastic Gradient Descent. Extensive evaluations on diverse benchmark datasets and neural network structures show that CORR-Reg achieves significant improvement over stateof-the-art regularization methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 25fca143-95cf-4bc8-a591-324bcbdd59b6Cited by top-tier papers4
- Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly DetectionJiaqi Zhu, Shaofeng Cai, Fang Deng, Beng Chin Ooi et al.ACM MM 2024 · 30 citations
- Robust and Transferable Log-based Anomaly DetectionPeng Jia, Shaofeng Cai, Beng Chin Ooi, Pinghui Wang et al.SIGMOD 2023 · 26 citations
- Database Native Model Selection: Harnessing Deep Neural Networks in Database SystemsNaili Xing, Shaofeng Cai, Gang Chen, Zhaojing Luo et al.VLDB 2024 · 14 citations
- Detecting Data Deviations in Electronic Health RecordsKaiping Zheng, Horng Ruey Chua, Beng Chin OoiNeurIPS 2025 · 1 citation
Related papers
- Regularized Pairwise Relationship based Analytics for Structured DataZhaojing Luo, Shaofeng Cai, Yatong Wang, Beng Chin OoiSIGMOD 2023 · 12 citations
- Matching Learned Causal Effects of Neural Networks with Domain PriorsSai Srinivas Kancheti, Abbavaram Gowtham Reddy, Vineeth N. Balasubramanian, Amit SharmaICML 2022 · 17 citations
- Pruning neural network models for gene regulatory dynamics using data and domain knowledgeIntekhab Hossain, Jonas Fischer, Rebekka Burkholz, John QuackenbushNeurIPS 2024 · 1 citation
- Learning from Failure: De-biasing Classifier from Biased ClassifierJun Hyun Nam, Hyuntak Cha, Sungsoo Ahn, Jaeho Lee et al.NeurIPS 2020 · 428 citations
- Visual Neural Decomposition to Explain Multivariate Data SetsJohannes Knittel, Andrés Lalama, Steffen Koch, Thomas ErtlIEEE VIS 2020 · 14 citations
