ColorFLUX: A Structure-Color Decoupling Framework for Old Photo Colorization
Bingchen Li, Zhixin Wang, Fan Li, Jiaqi Xu, Jiaming Guo, Renjing Pei, Xin Li, Zhibo Chen
Abstract
Old photos preserve invaluable historical memories, making their restoration and colorization highly desirable. While existing restoration models can address some degradation issues like denoising and scratch removal, they often struggle with accurate colorization. This limitation arises from the unique degradation inherent in old photos, such as faded brightness and altered color hues, which are different from modern photo distributions, creating a substantial domain gap during colorization. In this paper, we propose a novel old photo colorization framework based on the generative diffusion model FLUX. Our approach introduces a structure-color decoupling strategy that separates structure preservation from color restoration, enabling accurate colorization of old photos while maintaining structural consistency. We further enhance the model with a progressive Direct Preference Optimization (Pro-DPO) strategy, which allows the model to learn subtle color preferences through coarse-to-fine transitions in color augmentation. Additionally, we address the limitations of text-based prompts by introducing visual semantic prompts, which extract fine-grained semantic information directly from old photos, helping to eliminate the color bias inherent in old photos. Experimental results on both synthetic and real datasets demonstrate that our approach outperforms existing state-of-the-art colorization methods, including closed-source commercial models, producing high-quality and vivid colorization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5daef2c4-c4fe-470f-a696-134e52759014Builds on31
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
- Sigmoid Loss for Language Image Pre-TrainingXiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, Lucas BeyerICCV 2023 · 2,932 citations
- Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image GenerationYuval Kirstain, Adam Polyak, Uriel Singer, Shahbuland Matiana et al.NeurIPS 2023 · 1,192 citations
Related papers
- Self-Supervised Selective-Guided Diffusion Model for Old-Photo Face RestorationWenjie Li, Xiangyi Wang, Heng Guo, Guangwei Gao et al.NeurIPS 2025 · 13 citations
- CG-GAN: Class-Attribute Guided Generative Adversarial Network for Old Photo RestorationJixin Liu, Rui Chen, Shipeng An, Heng ZhangACM MM 2021 · 11 citations
- COCO-LC: Colorfulness Controllable Language-based ColorizationYifan Li, Yuhang Bai, Shuai Yang, Jiaying LiuACM MM 2024 · 7 citations
- PGDiff: Guiding Diffusion Models for Versatile Face Restoration via Partial GuidancePeiqing Yang, Shangchen Zhou, Qingyi Tao, Chen Change LoyNeurIPS 2023 · 82 citations
- Exploring Palette based Color Guidance in Diffusion ModelsQianru Qiu, Jiafeng Mao, Xueting WangACM MM 2025 · 4 citations
