L-CoIns: Language-based Colorization With Instance Awareness
Zheng Chang, Shuchen Weng, Peixuan Zhang, Yu Li, Si Li, Boxin Shi
摘要
The right woman is dressed in shirt of violet color. L-CoDe L-CoDer Ours ML2018 The middle woman is dressed in orange and the right woman in yellow. L-CoDe L-CoDer Ours ML2018 Three women are dressed in pink. L-CoDe L-CoDer Ours ML2018 The middle woman is dressed in yellow, the left woman in blue, the right woman in red. L-CoDe L-CoDer Ours ML2018 Grayscale Figure 1. Language-based colorization results given four different language descriptions, compared with ML2018 [29], L-CoDe [42], and L-CoDer [6]. Top left: For the description that has clear correspondences between color words and object words, our method correctly colorizes all corresponding regions. Top right: For the description that assigns distinct colors for every instance corresponding to the same object words, our model predicts the exact correspondence between the instance region and the color word. Botton left: For the description that includes unobserved correspondences between color words and object words, our method could adaptively parse the sentence and determine the correct semantics for colorization. Bottom right: For the description that is against the statistical correlation between luminance and color words, our method shows the robustness and colorize description-consistent results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- L-CAD: Language-based Colorization with Any-level Descriptions using Diffusion PriorsZheng Chang, Shuchen Weng, Peixuan Zhang, Yu Li 等NeurIPS 2023 · 被引用 42 次
- Affective Image Filter: Reflecting Emotions from Text to ImagesShuchen Weng, Peixuan Zhang, Zheng Chang, Xinlong Wang 等ICCV 2023 · 被引用 25 次
- Language-guided Image Reflection SeparationHaofeng Zhong, Yuchen Hong, Shuchen Weng, Jinxiu Liang 等CVPR 2024 · 被引用 14 次
- LuminAIRe: Illumination-Aware Conditional Image Repainting for Lighting-Realistic GenerationJiajun Tang, Haofeng Zhong, Shuchen Weng, Boxin ShiNeurIPS 2023 · 被引用 6 次
- Pre-training LiDAR-based 3D Object Detectors through ColorizationTai-Yu Pan, Chenyang Ma, Tianle Chen, Cheng Perng Phoo 等ICLR 2024 · 被引用 5 次
它引用的顶会 Paper18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- ViLT: Vision-and-Language Transformer Without Convolution or Region SupervisionWonjae Kim, Bokyung Son, Ildoo KimICML 2021 · 被引用 2,258 次
- MDETR - Modulated Detection for End-to-End Multi-Modal UnderstandingAishwarya Kamath, Mannat Singh, Yann LeCun, Gabriel Synnaeve 等ICCV 2021 · 被引用 1,114 次
相关 Paper
- L-CoDe: Language-Based Colorization Using Color-Object Decoupled ConditionsShuchen Weng, Hao Wu, Zheng Chang, Jiajun Tang 等AAAI 2022 · 被引用 57 次
- COCO-LC: Colorfulness Controllable Language-based ColorizationYifan Li, Yuhang Bai, Shuai Yang, Jiaying LiuACM MM 2024 · 被引用 7 次
- Dual Coding Theory in Action: Language-Assisted Human Pose Estimation in VideosSifan Wu, Haipeng Chen, Yingda Lyu, Shaojing Fan 等AAAI 2026
- Language-based Photo Color Adjustment for Graphic DesignsZhenwei Wang, Nanxuan Zhao, Gerhard P. Hancke, Rynson W. H. LauSIGGRAPH 2023 · 被引用 49 次
- Universal 3D Shape Matching via Coarse-to-Fine Language GuidanceQinfeng Xiao, Guofeng Mei, Bo Yang, Zhang Liying 等CVPR 2026 · 被引用 1 次
