Towards Photorealistic Video Colorization via Gated Color-Guided Image Diffusion Models
Jiaxing Li, Hongbo Zhao, Yijun Wang, Jianxin Lin
Abstract
Video colorization poses challenging tasks, necessitating structural stability, continuity, and details control in the colors produced. In this paper, based on a pretrained text-to-image model, we introduce the Gated Color Guidance module (GCG ), enabling the model to adaptively perform color propagation or generation according to the structural differences between reference and grayscale frames. Based on this multifunctionality, we propose a novel two-stage coloring strategy. In the first stage, under reference-mask condition, the model autonomously and jointly colors input keyframes in a one-to-many color domain mapping, while temporal coherence constraints are emphasized by modifying the attention mechanism. In the second stage, under reference-guided condition, the model effectively captures the colors of matching structures in the reference, and we further introduce Sliding Reference Grid strategy (SRG) to merge and extract the color features from multiple frames, providing more stable coloring for the grayscale frames. Through this pipeline, we can achieve high-quality and stable video coloring while maintaining the accuracy of detailed colors. Additionally, the two-stage strategy is flexible and detachable, allowing users to adjust the number of selected reference frames to balance coloring quality and efficiency. Extensive experiments demonstrate that our method significantly outperforms previous state-of-the-art models in both qualitative comparison and quantitative measurement.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 53141337-0b93-48af-8388-778dc331f8a6Cited by top-tier papers2
- Color3D: Controllable and Consistent 3D Colorization with Personalized ColorizerYecong Wan, Mingwen Shao, Renlong Wu, Wangmeng ZuoICLR 2026 · 4 citations
- FaceShot: Bring Any Character into LifeJunyao Gao, Yanan Sun, Fei Shen, Xin Jiang et al.ICLR 2025
Related papers
- ColorDiffuser: Video Colorization with Pretrained Text-to-Image Diffusion ModelsHanyuan Liu, Minshan Xie, Jinbo Xing, Chengze Li et al.ACM MM 2025 · 2 citations
- Versatile Vision Foundation Model for Image and Video ColorizationVukasin Bozic, Abdelaziz Djelouah, Yang Zhang, Radu Timofte et al.SIGGRAPH 2024 · 9 citations
- Gray2ColorNet: Transfer More Colors from Reference ImagePeng Lu, Jinbei Yu, Xujun Peng, Zhaoran Zhao et al.ACM MM 2020 · 55 citations
- Automatic Controllable Colorization via ImaginationXiaoyan Cong, Yue Wu, Qifeng Chen, Chenyang LeiCVPR 2024
- COCO-LC: Colorfulness Controllable Language-based ColorizationYifan Li, Yuhang Bai, Shuai Yang, Jiaying LiuACM MM 2024 · 7 citations
