Colorization Transformer
Manoj Kumar, Dirk Weissenborn, Nal Kalchbrenner
Abstract
We present the Colorization Transformer, a novel approach for diverse high fidelity image colorization based on self-attention. Given a grayscale image, the colorization proceeds in three steps. We first use a conditional autoregressive transformer to produce a low resolution coarse coloring of the grayscale image. Our architecture adopts conditional transformer layers to effectively condition grayscale input. Two subsequent fully parallel networks upsample the coarse colored low resolution image into a finely colored high resolution image. Sampling from the Colorization Transformer produces diverse colorings whose fidelity outperforms the previous state-of-the-art on colorising ImageNet based on FID results and based on a human evaluation in a Mechanical Turk test. Remarkably, in more than 60% of cases human evaluators prefer the highest rated among three generated colorings over the ground truth. The code and pre-trained checkpoints for Colorization Transformer are publicly available at https://github.com/google-research/google-research/tree/master/coltran
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69be0ccb-b336-4c4f-8526-1b711d397262Cited by top-tier papers15
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Palette: Image-to-Image Diffusion ModelsChitwan Saharia, William Chan, Huiwen Chang, Chris A. Lee et al.SIGGRAPH 2022 · 1,638 citations
- Towards Vivid and Diverse Image Colorization with Generative Color PriorYanze Wu, Xintao Wang, Yu Li, Honglun Zhang et al.ICCV 2021 · 123 citations
- PromptRestorer: A Prompting Image Restoration Method with Degradation PerceptionCong Wang, Jinshan Pan, Wei Wang, Jiangxin Dong et al.NeurIPS 2023 · 109 citations
- UViM: A Unified Modeling Approach for Vision with Learned Guiding CodesAlexander Kolesnikov, André Susano Pinto, Lucas Beyer, Xiaohua Zhai et al.NeurIPS 2022 · 88 citations
Builds on1
Related papers
- Versatile Vision Foundation Model for Image and Video ColorizationVukasin Bozic, Abdelaziz Djelouah, Yang Zhang, Radu Timofte et al.SIGGRAPH 2024 · 9 citations
- Automatic Controllable Colorization via ImaginationXiaoyan Cong, Yue Wu, Qifeng Chen, Chenyang LeiCVPR 2024
- DDColor: Towards Photo-Realistic Image Colorization via Dual DecodersXiaoyang Kang, Tao Yang, Wenqi Ouyang, Peiran Ren et al.ICCV 2023 · 84 citations
- Generating images with sparse representationsCharlie Nash, Jacob Menick, Sander Dieleman, Peter W. BattagliaICML 2021 · 291 citations
- Stylization-Based Architecture for Fast Deep Exemplar ColorizationZhongyou Xu, Tingting Wang, Faming Fang, Yun Sheng et al.CVPR 2020
