Text2Weight: Bridging Natural Language and Neural Network Weight Spaces
Bowen Tian, Wenshuo Chen, Zexi Li, Songning Lai, Jiemin Wu, Yutao Yue
摘要
How far are we really from automatically generating neural networks? While neural network weight generation shows promise, current approaches struggle with generalization to unseen tasks and practical application exploration. To address this, we propose T2W, a diffusion transformer framework that generates task-specific weights conditioned on natural language descriptions. T2W hierarchically processes network parameters into uniform blocks, integrates text embeddings from CLIP via a prior attention mechanism, and employs adversarial training with weight-space augmentation to enhance generalization. Experiments on Cifar100, Caltech256, and TinyImageNet demonstrate T2W's ability to produce high-quality weights for unseen tasks, outperforming optimization-based initialization and enabling novel applications such as weight enhancement and text-guided model fusion. Our work bridges textual semantics with weight-space dynamics, supported by an open-source dataset of text-weight pairs, advancing the practicality of generative models in neural network parameter synthesis. Our code is available on https://github.com/TianSuya/T2W.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Generative Adaptation of Dynamics to Environmental Shifts via Weight-space DiffusionRuikun Li, Huandong Wang, Jingtao Ding, Yuan Yuan 等ICML 2026 · 被引用 4 次
- WeightCLIP: Aligning Datasets and Models for Weight Space LearningAron Asefaw, Konstantinos Tzevelekakis, Damian Falk, Léo Meynent 等ICML 2026
- AnyEdit++: Adaptive Long-Form Knowledge Editing via Bayesian SurpriseBowen Tian, Caixue He, Jiemin Wu, Jingying Wang 等ICML 2026
它引用的顶会 Paper13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Equivariant Architectures for Learning in Deep Weight SpacesAviv Navon, Aviv Shamsian, Idan Achituve, Ethan Fetaya 等ICML 2023 · 被引用 101 次
- Permutation Equivariant Neural FunctionalsAllan Zhou, Kaien Yang, Kaylee Burns, Adriano Cardace 等NeurIPS 2023 · 被引用 84 次
相关 Paper
- NNiT: Width-Agnostic Neural Network Generation with Structurally Aligned Weight SpacesJiwoo Kim, Swarajh Mehta, Hao-Lun Hsu, Hyunwoo Ryu 等ICML 2026
- Learning to Imagine: Visually-Augmented Natural Language GenerationTianyi Tang, Yushuo Chen, Yifan Du, Junyi Li 等ACL 2023 · 被引用 7 次
- Condition-Aware Neural Network for Controlled Image GenerationHan Cai, Muyang Li, Qinsheng Zhang, Ming-Yu Liu 等CVPR 2024
- CLIPTexture: Text-Driven Texture SynthesisYiren SongACM MM 2022 · 被引用 7 次
- Learning Dynamic Prior Knowledge for Text-to-Face Pixel SynthesisJun Peng, Xiaoxiong Du, Yiyi Zhou, Jing He 等ACM MM 2022 · 被引用 6 次
