Multimodal Color Recommendation in Vector Graphic Documents
Qianru Qiu, Xueting Wang, Mayu Otani
摘要
Color selection plays a critical role in graphic document design and requires sufficient consideration of various contexts. However, recommending appropriate colors which harmonize with the other colors and textual contexts in documents is a challenging task, even for experienced designers. In this study, we propose a multimodal masked color model that integrates both color and textual contexts to provide text-aware color recommendation for graphic documents. Our proposed model comprises self-attention networks to capture the relationships between colors in multiple palettes, and cross-attention networks that incorporate both color and CLIP-based text representations. Our proposed method primarily focuses on color palette completion, which recommends colors based on the given colors and text. Additionally, it is applicable for another color recommendation task, full palette generation, which generates a complete color palette corresponding to the given text. Experimental results demonstrate that our proposed approach surpasses previous color palette completion methods on accuracy, color distribution, and user experience, as well as full palette generation methods concerning color diversity and similarity to the ground truth palettes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Exploring Palette based Color Guidance in Diffusion ModelsQianru Qiu, Jiafeng Mao, Xueting WangACM MM 2025 · 被引用 4 次
- Multimodal Markup Document Models for Graphic Design CompletionKotaro Kikuchi, Ukyo Honda, Naoto Inoue, Mayu Otani 等ACM MM 2025 · 被引用 1 次
它引用的顶会 Paper3
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- CanvasVAE: Learning to Generate Vector Graphic DocumentsKota YamaguchiICCV 2021 · 被引用 103 次
- Dynamic closest color warping to sort and compare palettesSuzi Kim, Sunghee ChoiSIGGRAPH 2021 · 被引用 64 次
相关 Paper
- LapsCore: Language-guided Person Search via Color ReasoningYushuang Wu, Zizheng Yan, Xiaoguang Han, Guanbin Li 等ICCV 2021 · 被引用 91 次
- Text-Guided Image InpaintingZijian Zhang, Zhou Zhao, Zhu Zhang, Baoxing Huai 等ACM MM 2020 · 被引用 17 次
- Mask to Reconstruct: Cooperative Semantics Completion for Video-text RetrievalHan Fang, Zhifei Yang, Xianghao Zang, Chao Ban 等ACM MM 2023 · 被引用 4 次
- ARMANI: Part-level Garment-Text Alignment for Unified Cross-Modal Fashion DesignXujie Zhang, Yu Sha, Michael C. Kampffmeyer, Zhenyu Xie 等ACM MM 2022 · 被引用 27 次
- L-CAD: Language-based Colorization with Any-level Descriptions using Diffusion PriorsZheng Chang, Shuchen Weng, Peixuan Zhang, Yu Li 等NeurIPS 2023 · 被引用 42 次
