GenColor: Generative and Expressive Color Enhancement with Pixel-Perfect Texture Preservation
Yi Dong, Yuxi Wang, Xianhui Lin, Wenqi Ouyang, Zhiqi Shen, Peiran Ren, Ruoxi Fan, Rynson W. H. Lau
摘要
Generative methods ( e.g ., Midjourney), on the other hand, create dramatic transformations but alter textures, compromising authenticity. The difference maps ( 2 nd row) show GenColor making selective region-specific adjustments similar to Midjourney, but with better results than Expert C (the best retoucher in Adobe5K [2]), which is limited to global filter adjustments. Abstract Color enhancement is a crucial yet challenging task in digital photography. It demands methods that are (i) expressive enough for fine-grained adjustments, (ii) adaptable to diverse inputs, and (iii) able to preserve texture. Existing approaches typically fall short in at least one of these aspects, yielding unsatisfactory re-sults. We propose GenColor, a novel diffusion-based framework for sophisticated, texture-preserving color enhancement. GenColor reframes the task as conditional image generation. Leveraging ControlNet and a tailored training scheme, it learns advanced color transformations that adapt to diverse lighting and content. We train GenColor on ARTISAN, our newly collected large-scale
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 被引用 6,549 次
- Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined LevelsHaoning Wu, Zicheng Zhang, Weixia Zhang, Chaofeng Chen 等ICML 2024 · 被引用 499 次
相关 Paper
- Versatile Vision Foundation Model for Image and Video ColorizationVukasin Bozic, Abdelaziz Djelouah, Yang Zhang, Radu Timofte 等SIGGRAPH 2024 · 被引用 9 次
- CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image GenerationKangfu Mei, Mauricio Delbracio, Hossein Talebi, Zhengzhong Tu 等CVPR 2024 · 被引用 11 次
- ChromaFusionNet (CFNet): Natural Fusion of Fine-Grained Color EditingYi Dong, Yuxi Wang, Ruoxi Fan, Wenqi Ouyang 等AAAI 2024 · 被引用 2 次
- Unpaired Image Enhancement Featuring Reinforcement-Learning-Controlled Image Editing SoftwareSatoshi Kosugi, Toshihiko YamasakiAAAI 2020 · 被引用 104 次
- IntrinsicEdit: Precise generative image manipulation in intrinsic spaceLinjie Lyu, Valentin Deschaintre, Yannick Hold-Geoffroy, Milos Hasan 等SIGGRAPH 2025 · 被引用 7 次
