A Study on PAVE Specification for Learnware
Hao-Yu Shi, Zhi-Hao Tan, Zi-Chen Zhao, Yang Yu, Zhi-Hua Zhou
摘要
``Learnware = Model + Specification''. A learnware comprises a submitted model paired with a specification sketching its capabilities. For a Learnware Dock System (LDS) which accommodates numerous models, these specifications are essential to enabling users to identify helpful models, eliminating the requirement for prohibitively costly per-model evaluations. Recently, Parameter Vector (PAVE) specification, which utilizes the changes in pre-trained model parameters to inherently encode the model capability and task requirements, shows promising capabilities in enabling identifying useful learnwares for high-dimensional, unstructured text data. In this paper, we present a comprehensive study of PAVE specification for learnware identification. Theoretically, from the neural tangent kernel perspective, we establish a tight connection between PAVE and prior specifications, providing a theoretical explanation for their shared underlying principles. We further approximate PAVE in a low-rank space and analyze the approximation error bound, highly reducing the computational and storage overhead. Extensive empirical studies demonstrate that PAVE specification excels at identifying CV and NLP learnwares even from heterogeneous learnware repository with corrupted model quality. Reusing identified learnware to solve user tasks can even outperform user-fine-tuned pre-trained models in data-limited scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Constructive Specification for Plug-and-Play Learnware AgentsJian-Dong Liu, Zi-Chen Zhao, Hao Sun, Lin-Xing Wu 等KDD 2026 · 被引用 3 次
- UNITE: Universal kNowledge Integration from Task-specific ExpertsShuxia Lin, Qiufeng Wang 00002, Xu Yang, Xin GengICLR 2026
- Identifying Learnwares via Reduced Neural Conditional Mean EmbeddingZi-Yu Mao, Ming LiICML 2026
它引用的顶会 Paper30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIsYujia Qin, Shihao Liang, Yining Ye, Kunlun Zhu 等ICLR 2024 · 被引用 1,469 次
- TIES-Merging: Resolving Interference When Merging ModelsPrateek Yadav, Derek Tam, Leshem Choshen, Colin A. Raffel 等NeurIPS 2023 · 被引用 999 次
- Dataset Condensation with Gradient MatchingBo Zhao, Konda Reddy Mopuri, Hakan BilenICLR 2021 · 被引用 684 次
相关 Paper
- Learnware Specification via Label-Aware Neural EmbeddingWei Chen, Junxiang Mao, Min-Ling ZhangAAAI 2025 · 被引用 1 次
- Handling Learnwares from Heterogeneous Feature Spaces with Explicit Label ExploitationPeng Tan, Hai-Tian Liu, Zhi-Hao Tan, Zhi-Hua ZhouNeurIPS 2024 · 被引用 8 次
- Integrated Learnware Identification and Reuse via Reusability-Aware Metric LearningHai-Tian Liu, Peng Tan, Jian-Dong Liu, Zhi-Hao Tan 等KDD 2026
- Learnware Specification via Dual AlignmentWei Chen, Junxiang Mao, Xiaozheng Wang, Min-Ling ZhangICML 2025
- On the Ability of Developers' Training Data Preservation of LearnwareHao-Yi Lei, Zhi-Hao Tan, Zhi-Hua ZhouNeurIPS 2024 · 被引用 10 次
