TESS 2: A Large-Scale Generalist Diffusion Language Model
Jaesung Tae, Hamish Ivison, Sachin Kumar, Arman Cohan
Abstract
We introduce TESS 2, a general instruction-following diffusion language model that outperforms contemporary instruction-tuned diffusion models, as well as matches and sometimes exceeds strong autoregressive (AR) models. We train TESS 2 by first adapting a strong AR model via continued pretraining with the usual cross-entropy as diffusion loss, and then performing further instruction tuning. We find that adaptation training as well as the choice of the base model is crucial for training good instruction-following diffusion models. We further propose reward guidance, a novel and modular inference-time guidance procedure to align model outputs without needing to train the underlying model. Finally, we show that TESS 2 further improves with increased inference-time compute, highlighting the utility of diffusion LMs in having fine-grained controllability over the amount of compute used at inference time. Code and models are available at https://github.com/hamishivi/tess-2.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b0aeae9d-ab8f-4344-bd37-ca554b2ff67bCited by top-tier papers5
- Continuously Augmented Discrete Diffusion model for Categorical Generative ModelingHuangjie Zheng, Shansan Gong, Ruixiang Zhang, Tianrong Chen et al.ICLR 2026 · 30 citations
- Non-Markovian Discrete Diffusion with Causal Language ModelsYangtian Zhang, Sizhuang He, Daniel LeVine, Lawrence Zhao et al.NeurIPS 2025 · 5 citations
- Don't Let It Fade: Preserving Edits in Diffusion Language Models via Token Timestep AllocationWoojin Kim, Jaeyoung DoNeurIPS 2025 · 2 citations
- Unlocking the Potential of Diffusion Language Models through Template InfillingJunhoo Lee, Seungyeon Kim, Nojun KwakACL 2026 · 1 citation
- Efficient Self-Evaluation for Diffusion Language Models via Sequence RegenerationLinhao Zhong, Linyu Wu, Wen Wang, Yuling Xi et al.ACL 2026
Builds on27
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
Related papers
- Reward-Instruct: A Reward-Centric Approach to Fast Photo-Realistic Image GenerationYihong Luo, Tianyang Hu, Weijian Luo, Kenji Kawaguchi et al.NeurIPS 2025 · 20 citations
- Large Language Models to Diffusion FinetuningEdoardo Cetin, Tianyu Zhao, Yujin TangICML 2025
- RewardFlow: Generate Images by Optimizing What You RewardOnkar Susladkar, Dong-Hwan Jang, Tushar Prakash, Adheesh Sunil Juvekar et al.CVPR 2026 · 2 citations
- InstructVideo: Instructing Video Diffusion Models with Human FeedbackHangjie Yuan, Shiwei Zhang, Xiang Wang, Yujie Wei et al.CVPR 2024 · 13 citations
- PromptLoop: Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model AlignmentSuhyeon Lee, Jong Chul YeCVPR 2026
