One More Step: A Versatile Plug-and-Play Module for Rectifying Diffusion Schedule Flaws and Enhancing Low-Frequency Controls
Minghui Hu, Jianbin Zheng, Chuanxia Zheng, Chaoyue Wang, Dacheng Tao, Tat-Jen Cham
2024Year
6Top-tier citations
Abstract
Close-up portrait of a man wearing suit posing in a dark studio, rim lighting, teal hue, octane, unreal A bald eagle against a white background A dark town square lit only by a few torchlights Solid black background Majestic white angel sculpture inside a solemn cathedral, exquisite details and a sacred ambiance, bright and pure visual effects, extreme white background Prompt (a) (b)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Phased Consistency ModelsFu-Yun Wang, Zhaoyang Huang, Alexander William Bergman, Dazhong Shen et al.NeurIPS 2024 · 86 citations
- Golden Noise for Diffusion Models: A Learning FrameworkZikai Zhou, Shitong Shao, Lichen Bai, Shufei Zhang et al.ICCV 2025 · 9 citations
- Redefining in Dictionary: Towards an Enhanced Semantic Understanding of Creative GenerationFu Feng, Yucheng Xie, Xu Yang, Jing Wang et al.CVPR 2025
- Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority GenerationSoobin Um, Beomsu Kim, Jong Chul YeICML 2025
- Rectifying the Emotional Flow: Aligning Priors and Dynamic Guidance for High-Arousal Text-to-SpeechFangming Feng, Dongjie Fu, Zequn Xie, Yu Zhang et al.ACL 2026
Builds on14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- LucidDreamer: Towards High-Fidelity Text-to-3D Generation via Interval Score MatchingYixun Liang, Xin Yang, Jiantao Lin, Haodong Li et al.CVPR 2024
- Mimir: Improving Video Diffusion Models for Precise Text UnderstandingShuai Tan, Biao Gong, Yutong Feng, Kecheng Zheng et al.CVPR 2025
- Text-to-3D Generation with Bidirectional Diffusion Using Both 2D and 3D PriorsLihe Ding, Shaocong Dong, Zhanpeng Huang, Zibin Wang et al.CVPR 2024
- PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention SteeringYibin Wang, Weizhong Zhang, Jianwei Zheng, Cheng JinACM MM 2024 · 9 citations
- Total relighting: learning to relight portraits for background replacementRohit Pandey, Sergio Orts-Escolano, Chloe LeGendre, Christian Häne et al.SIGGRAPH 2021 · 138 citations
