Understanding the Impact of Introducing Constraints at Inference Time on Generalization Error
Masaaki Nishino, Kengo Nakamura, Norihito Yasuda
摘要
Since machine learning technologies are being used in various practical situations, models with merely low prediction errors might not be satisfactory; prediction errors occurring with a low probability might yield dangerous results in some applications. Therefore, there are attempts to achieve an ML model whose input-output pairs are guaranteed to satisfy given constraints. Among such attempts, many previous works chose the approach of modifying the outputs of an ML model at the inference time to satisfy the constraints. Such a strategy is handy because we can control its output without expensive training or fine-tuning. However, it is unclear whether using constraints only in the inference time degrades a model's predictive performance. This paper analyses how the generalization error bounds change when we only put constraints in the inference time. Our main finding is that a class of loss functions preserves the relative generalization error, i.e., the difference in generalization error compared with the best model will not increase by imposing constraints at the inference time on multi-class classification. Some popular loss functions preserve the relative error, including the softmax cross-entropy loss. On the other hand, we also show that some loss functions do not preserve relative error when we use constraints. Our results suggest the importance of choosing a suitable loss function when we only use constraints in the inference time.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- COLD Decoding: Energy-based Constrained Text Generation with Langevin DynamicsLianhui Qin, Sean Welleck, Daniel Khashabi, Yejin ChoiNeurIPS 2022 · 被引用 217 次
- Semantic Probabilistic Layers for Neuro-Symbolic LearningKareem Ahmed, Stefano Teso, Kai-Wei Chang, Guy Van den Broeck 等NeurIPS 2022 · 被引用 133 次
- MultiplexNet: Towards Fully Satisfied Logical Constraints in Neural NetworksNick Hoernle, Rafael-Michael Karampatsis, Vaishak Belle, Kobi GalAAAI 2022 · 被引用 73 次
- Tractable Control for Autoregressive Language GenerationHonghua Zhang, Meihua Dang, Nanyun Peng, Guy Van den BroeckICML 2023 · 被引用 63 次
- Learning with Explanation ConstraintsRattana Pukdee, Dylan Sam, J. Zico Kolter, Maria-Florina Balcan 等NeurIPS 2023 · 被引用 11 次
相关 Paper
- Generalization Analysis on Learning with a Concurrent VerifierMasaaki Nishino, Kengo Nakamura, Norihito YasudaNeurIPS 2022 · 被引用 1 次
- On Regularization and Inference with Label ConstraintsKaifu Wang, Hangfeng He, Tin D. Nguyen, Piyush Kumar 等ICML 2023 · 被引用 6 次
- The Devil is in the Margin: Margin-based Label Smoothing for Network CalibrationBingyuan Liu, Ismail Ben Ayed, Adrian Galdran, Jose DolzCVPR 2022 · 被引用 61 次
- DeepSaDe: Learning Neural Networks That Guarantee Domain Constraint SatisfactionKshitij Goyal, Sebastijan Dumancic, Hendrik BlockeelAAAI 2024 · 被引用 9 次
- Learning where and when to reason in neuro-symbolic inferenceCristina Cornelio, Jan Stuehmer, Shell Xu Hu, Timothy M. HospedalesICLR 2023
