Generalization Analysis for Controllable Learning
Yifan Zhang, Xiao Zhang, Min-Ling Zhang
Abstract
Controllability has become a critical issue in trustworthy machine learning, as a controllable learner allows for dynamic model adaptation to task requirements during testing. However, existing research lacks a comprehensive understanding of how to effectively measure and analyze the generalization performance of controllable learning methods. In an attempt to move towards this goal from a generalization perspective, we first establish a unified framework for controllable learning. Then, we develop a novel vector-contraction inequality and derive a tight generalization bound for general controllable learning classes, which is independent of the number of task targets except for logarithmic factors and represents the current best-in-class theoretical result. Furthermore, we derive generalization bounds for two typical controllable learning methods: embedding-based and hypernetwork-based methods. We also upper bound the Rademacher complexities of commonly used control and prediction functions, which serve as modular theoretical components for deriving generalization bounds for specific controllable learning methods in practical applications such as recommender systems. Our theoretical results without strong assumptions provide general theoretical guarantees for controllable learning methods and offer new insights into understanding controllability in machine learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on15
- Inductive Biases and Variable Creation in Self-Attention MechanismsBenjamin L. Edelman, Surbhi Goel, Sham M. Kakade, Cyril ZhangICML 2022 · 154 citations
- HyperPrompt: Prompt-based Task-Conditioning of TransformersYun He, Huaixiu Steven Zheng, Yi Tay, Jai Prakash Gupta et al.ICML 2022 · 110 citations
- Aligning Large Language Models with Representation Editing: A Control PerspectiveLingkai Kong, Haorui Wang, Wenhao Mu, Yuanqi Du et al.NeurIPS 2024 · 80 citations
- On the Modularity of HypernetworksTomer Galanti, Lior WolfNeurIPS 2020 · 79 citations
- APG: Adaptive Parameter Generation Network for Click-Through Rate PredictionBencheng Yan, Pengjie Wang, Kai Zhang, Feng Li et al.NeurIPS 2022 · 36 citations
Related papers
- Tight and Fast Bounds for Multi-Label LearningYifan Zhang, Min-Ling ZhangICML 2025
- A Non-Asymptotic Moreau Envelope Theory for High-Dimensional Generalized Linear ModelsLijia Zhou, Frederic Koehler, Pragya Sur, Danica J. Sutherland et al.NeurIPS 2022 · 13 citations
- Generalization Analysis on Learning with a Concurrent VerifierMasaaki Nishino, Kengo Nakamura, Norihito YasudaNeurIPS 2022 · 1 citation
- Nearly-tight Bounds for Deep Kernel LearningYifan Zhang, Min-Ling ZhangICML 2023 · 3 citations
- Optimistic Bounds for Multi-output LearningHenry W. J. Reeve, Ata KabánICML 2020 · 14 citations
