Generation with Dynamic Vocabulary
Yanting Liu, Tao Ji, Changzhi Sun, Yuanbin Wu, Xiaoling Wang
摘要
We introduce a new dynamic vocabulary for language models. It can involve arbitrary text spans during generation. These text spans act as basic generation bricks, akin to tokens in the traditional static vocabularies. We show that, the ability to generate multi-tokens atomically improve both generation quality and efficiency (compared to the standard language model, the MAUVE metric is increased by 25%, the latency is decreased by 20%). The dynamic vocabulary can be deployed in a plug-and-play way, thus is attractive for various downstream applications. For example, we demonstrate that dynamic vocabulary can be applied to different domains in a training-free manner. It also helps to generate reliable citations in question answering tasks (substantially enhancing citation results without compromising answer accuracy). 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Protein Design with Dynamic Protein VocabularyNuowei Liu, Jiahao Kuang, Yanting Liu, Tao Ji 等NeurIPS 2025 · 被引用 12 次
- Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language ModelsTobias Schreieder, Tim Schopf, Michael FärberACL 2026 · 被引用 10 次
- zip2zip: Inference-Time Adaptive Tokenization via Online CompressionSaibo Geng, Nathan Ranchin, Yunzhen Yao, Maxime Peyrard 等NeurIPS 2025 · 被引用 5 次
- On Discriminative vs. Generative classifiers: Rethinking MLLMs for Action UnderstandingZhanzhong Pang, Dibyadip Chatterjee, Fadime Sener, Angela YaoICLR 2026 · 被引用 1 次
- Flow of Spans: Generalizing Language Models to Dynamic Span-Vocabulary via GFlowNetsBo Xue, Yunchong Song, Fanghao Shao, Xuekai Zhu 等ICLR 2026
它引用的顶会 Paper22
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai 等ICML 2022 · 被引用 1,629 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Generalization through Memorization: Nearest Neighbor Language ModelsUrvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer 等ICLR 2020 · 被引用 1,038 次
相关 Paper
- Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language ModelsYukun Huang, Sanxing Chen, Jian Pei, Manzil Zaheer 等ICLR 2026 · 被引用 1 次
- DyVo: Dynamic Vocabularies for Learned Sparse Retrieval with EntitiesThong Nguyen, Shubham Chatterjee, Sean MacAvaney, Iain Mackie 等EMNLP 2024 · 被引用 4 次
- Retrieval is Accurate GenerationBowen Cao, Deng Cai, Leyang Cui, Xuxin Cheng 等ICLR 2024 · 被引用 13 次
- Copy is All You NeedTian Lan, Deng Cai, Yan Wang, Heyan Huang 等ICLR 2023
- Forking Paths in Neural Text GenerationEric J. Bigelow, Ari Holtzman, Hidenori Tanaka, Tomer David UllmanICLR 2025
