Guided Score identity Distillation for Data-Free One-Step Text-to-Image Generation
Mingyuan Zhou, Zhendong Wang, Huangjie Zheng, Hai Huang
Abstract
Google DeepMind, and 3 Atlassian distinguished older gentleman in a vintage study, surrounded by books and dim lighting [...] saharian landscape at sunset , 4k ultra realism, BY Anton Gorlin, trending on artstation, [...] chinese red blouse, in the style of dreamy and romantic compositions, floral explosions. [...] futuristic simple multilayered architecture, habitation cabin in the trees, dramatic soft light [...] poster art for the collection of the asian woman, dynamic anime [...], mysterious realism, [...] A fantasy-themed portrait of a female elf with golden hair and violet eyes, her attire [...] very beautiful girl in [...], white short top, charismatic personality, professional photo, [...] steampunk atmosphere, a stunning girl with a mecha musume aesthetic, adorned in [...] (Pirate ship sailing into a bioluminescence sea with a galaxy in the sky), epic, 4k, ultra. tshirt design, colourful, no background, yoda with sun glasses, dancing at a festival [...] 8k.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0f1e9062-6695-4a7b-ba4c-c2e240c39149Cited by top-tier papers3
- OmiAD: One-Step Adaptive Masked Diffusion Model for Multi-class Anomaly Detection via Adversarial DistillationYaoxuan Feng, Wenchao Chen, Yuxin Li, Bo Chen et al.ICML 2025
- Restoring Initial Noise Sensitivity in Text-to-Image Distillation through Geometric Alignmenthuayang Huang, Ruoyu Wang, Jinhui Zhao, Wei Deng et al.ICML 2026
- Random Conditioning for Diffusion Model Compression with DistillationDohyun Kim, Sehwan Park, Geonhee Han, Seung Wook Kim et al.CVPR 2025
Builds on62
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- SnapGen: Taming High-Resolution Text-to-Image Models for Mobile Devices with Efficient Architectures and TrainingJierun Chen, Dongting Hu, Xijie Huang, Huseyin Coskun et al.CVPR 2025
- Align Your Latents: High-Resolution Video Synthesis with Latent Diffusion ModelsAndreas Blattmann, Robin Rombach, Huan Ling, Tim Dockhorn et al.CVPR 2023
- LucidDreamer: Towards High-Fidelity Text-to-3D Generation via Interval Score MatchingYixun Liang, Xin Yang, Jiantao Lin, Haodong Li et al.CVPR 2024
- GLIGEN: Open-Set Grounded Text-to-Image GenerationYuheng Li, Haotian Liu, Qingyang Wu, Fangzhou Mu et al.CVPR 2023
- LoRACLR: Contrastive Adaptation for Customization of Diffusion ModelsEnis Simsar, Thomas Hofmann, Federico Tombari, Pinar YanardagCVPR 2025
