Language to Network: Conditional Parameter Adaptation with Natural Language Descriptions
Tian Jin, Zhun Liu, Shengjia Yan, Alexandre E. Eichenberger, Louis-Philippe Morency
摘要
Transfer learning using ImageNet pre-trained models has been the de facto approach in a wide range of computer vision tasks. However, fine-tuning still requires task-specific training data. In this paper, we propose N 3 (Neural Networks from Natural Language) -a new paradigm of synthesizing task-specific neural networks from language descriptions and a generic pre-trained model. N 3 leverages language descriptions to generate parameter adaptations as well as a new task-specific classification layer for a pre-trained neural network, effectively "fine-tuning" the network for a new task using only language descriptions as input. To the best of our knowledge, N 3 is the first method to synthesize entire neural networks from natural language. Experimental results show that N 3 can out-perform previous natural-language based zero-shot learning methods across 4 different zero-shot image classification benchmarks. We also demonstrate a simple method to help identify keywords in language descriptions leveraged by N 3 when synthesizing model parameters. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Prototype-based HyperAdapter for Sample-Efficient Multi-task TuningHao Zhao, Jie Fu, Zhaofeng HeEMNLP 2023 · 被引用 3 次
- Parameter-efficient Multi-task Fine-tuning for Transformers via Shared HypernetworksRabeeh Karimi Mahabadi, Sebastian Ruder, Mostafa Dehghani, James HendersonACL 2021
相关 Paper
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image CollectionsMuhammad Jehanzeb Mirza, Leonid Karlinsky, Wei Lin, Horst Possegger 等NeurIPS 2023 · 被引用 63 次
- ImaginaryNet: Learning Object Detectors without Real Images and AnnotationsMinheng Ni, Zitong Huang, Kailai Feng, Wangmeng ZuoICLR 2023 · 被引用 5 次
- Revisiting Classifier: Transferring Vision-Language Models for Video RecognitionWenhao Wu, Zhun Sun, Wanli OuyangAAAI 2023 · 被引用 141 次
- Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot GeneralizationYang Shen, Xiu-Shen Wei, Yifan Sun, Yuxin Song 等ICML 2025
