Dream Video: Composing Your Dream Videos with Customized Subject and Motion
Yujie Wei, Shiwei Zhang, Zhiwu Qing, Hangjie Yuan, Zhiheng Liu, Yu Liu, Yingya Zhang, Jingren Zhou, Hongming Shan
2024Year
50Top-tier citations
Abstract
Customized generation using diffusion models has made impressive progress in image generation, but remains unsatisfactory in the challenging video generation task, as it
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4b8af460-75ec-4e74-b4fc-a88b887354edCited by top-tier papers50
- CustomCrafter: Customized Video Generation with Preserving Motion and Concept Composition AbilitiesTao Wu, Yong Zhang, Xintao Wang, Xianpan Zhou et al.AAAI 2025 · 62 citations
- Direct-a-Video: Customized Video Generation with User-Directed Camera Movement and Object MotionShiyuan Yang, Liang Hou, Haibin Huang, Chongyang Ma et al.SIGGRAPH 2024 · 46 citations
- MagCache: Fast Video Generation with Magnitude-Aware CacheZehong Ma, Longhui Wei, Feng Wang, Shiliang Zhang et al.NeurIPS 2025 · 41 citations
- Stand-In: A Lightweight and Plug-and-Play Identity Control for Video GenerationBowen Xue, Zheng-Peng Duan, Qixin Yan, Wenjing Wang et al.CVPR 2026 · 28 citations
- Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing GuidanceYujie Wei, Shiwei Zhang, Hangjie Yuan, Yujin Han et al.ICLR 2026 · 26 citations
Builds on49
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 11,724 citations
Related papers
- Efficient Denoising Diffusion via Probabilistic MaskingWeizhong Zhang, Zhiwei Zhang, Renjie Pi, Zhongming Jin et al.ICML 2024
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
- Otil: Accelerating Diffusion Model Inference via Communication-Efficient Multi-GPU ParallelismXin Li, Shujun Tian, Tao Lu, Han Bao et al.CVPR 2026
- FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less ComputeSotiris Anagnostidis, Gregor Bachmann, Yeongmin Kim, Jonas Kohler et al.CVPR 2025
- MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video GenerationLudan Ruan, Yiyang Ma, Huan Yang, Huiguo He et al.CVPR 2023
