Text2Weight: Bridging Natural Language and Neural Network Weight Spaces
Bowen Tian, Wenshuo Chen, Zexi Li, Songning Lai, Jiemin Wu, Yutao Yue
Abstract
How far are we really from automatically generating neural networks? While neural network weight generation shows promise, current approaches struggle with generalization to unseen tasks and practical application exploration. To address this, we propose T2W, a diffusion transformer framework that generates task-specific weights conditioned on natural language descriptions. T2W hierarchically processes network parameters into uniform blocks, integrates text embeddings from CLIP via a prior attention mechanism, and employs adversarial training with weight-space augmentation to enhance generalization. Experiments on Cifar100, Caltech256, and TinyImageNet demonstrate T2W's ability to produce high-quality weights for unseen tasks, outperforming optimization-based initialization and enabling novel applications such as weight enhancement and text-guided model fusion. Our work bridges textual semantics with weight-space dynamics, supported by an open-source dataset of text-weight pairs, advancing the practicality of generative models in neural network parameter synthesis. Our code is available on https://github.com/TianSuya/T2W.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b0cda063-3a46-4bb2-9343-739350a81f74Cited by top-tier papers3
- Generative Adaptation of Dynamics to Environmental Shifts via Weight-space DiffusionRuikun Li, Huandong Wang, Jingtao Ding, Yuan Yuan et al.ICML 2026 · 4 citations
- WeightCLIP: Aligning Datasets and Models for Weight Space LearningAron Asefaw, Konstantinos Tzevelekakis, Damian Falk, Léo Meynent et al.ICML 2026
- AnyEdit++: Adaptive Long-Form Knowledge Editing via Bayesian SurpriseBowen Tian, Caixue He, Jiemin Wu, Jingying Wang et al.ICML 2026
Builds on13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Equivariant Architectures for Learning in Deep Weight SpacesAviv Navon, Aviv Shamsian, Idan Achituve, Ethan Fetaya et al.ICML 2023 · 101 citations
- Permutation Equivariant Neural FunctionalsAllan Zhou, Kaien Yang, Kaylee Burns, Adriano Cardace et al.NeurIPS 2023 · 84 citations
Related papers
- NNiT: Width-Agnostic Neural Network Generation with Structurally Aligned Weight SpacesJiwoo Kim, Swarajh Mehta, Hao-Lun Hsu, Hyunwoo Ryu et al.ICML 2026
- Learning to Imagine: Visually-Augmented Natural Language GenerationTianyi Tang, Yushuo Chen, Yifan Du, Junyi Li et al.ACL 2023 · 7 citations
- Condition-Aware Neural Network for Controlled Image GenerationHan Cai, Muyang Li, Qinsheng Zhang, Ming-Yu Liu et al.CVPR 2024
- CLIPTexture: Text-Driven Texture SynthesisYiren SongACM MM 2022 · 7 citations
- Learning Dynamic Prior Knowledge for Text-to-Face Pixel SynthesisJun Peng, Xiaoxiong Du, Yiyi Zhou, Jing He et al.ACM MM 2022 · 6 citations
