Vector Graphics Generation via Mutually Impulsed Dual-Domain Diffusion
Zhongyin Zhao, Ye Chen, Zhangli Hu, Xuanhong Chen, Bingbing Ni
摘要
Intelligent generation of vector graphics has very promising applications in the fields of advertising and logo design, artistic painting, animation production, etc. However, current mainstream vector image generation methods lack the encoding of image appearance information that is associated with the original vector representation and therefore lose valid supervision signal from the strong correlation between the discrete vector parameter (drawing in-struction) sequence and the target shape/structure of the corresponding pixel image. On the one hand, the gener-ation process based on pure vector domain completely ignores the similarity measurement between shape parameter (and their combination) and the paired pixel image appearance pattern; on the other hand, two-stage methods (i.e., generation-and-vectorization) based on pixel diffusion followed by differentiable image-to-vector translation suf-fer from wrong error-correction signal caused by approxi-mate gradients. To address the above issues, we propose a novel generation framework based on dual-domain (vector-pixel) diffusion with cross-modality impulse signals from each other. First, in each diffusion step, the current representation extracted from the other domain is used as a condition variable to constrain the subsequent sampling operation, yielding shape-aware new parameterizations; second, independent supervision signals from both domains avoid the gradient error accumulation problem caused by cross-domain representation conversion. Extensive experimental results on popular benchmarks including font and icon datasets demonstrate the great advantages of our proposed framework in terms of generated shape quality.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- LottieGPT: Tokenizing Vector Animation for Autoregressive GenerationJunhao Chen, Kejun Gao, Yuehan Cui, Mingze Sun 等CVPR 2026 · 被引用 10 次
- SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG GenerationHanqi Chen, Zhongyin Zhao, Ye Chen, Zhujin Liang 等ACM MM 2025 · 被引用 3 次
- OmniLottie: Generating Vector Animations via Parameterized Lottie TokensYiying Yang, Wei Cheng, Sijin Chen, Honghao Fu 等CVPR 2026 · 被引用 2 次
- Easy-editable Image Vectorization with Multi-layer Multi-scale Distributed Visual Feature EmbeddingYe Chen, Zhangli Hu, Zhongyin Zhao, Yupeng Zhu 等CVPR 2025
它引用的顶会 Paper14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang 等NeurIPS 2022 · 被引用 1,546 次
相关 Paper
- NIVeL: Neural Implicit Vector Layers for Text-to-Vector GenerationVikas Thamizharasan, Difan Liu, Matthew Fisher, Nanxuan Zhao 等CVPR 2024 · 被引用 5 次
- Im2Vec: Synthesizing Vector Graphics Without Vector SupervisionPradyumna Reddy, Michaël Gharbi, Michal Lukác, Niloy J. MitraCVPR 2021
- Visual Layout Composer: Image-Vector Dual Diffusion Model for Design Layout GenerationMohammad Amin Shabani, Zhaowen Wang, Difan Liu, Nanxuan Zhao 等CVPR 2024 · 被引用 7 次
- Text-to-Vector Generation with Neural Path RepresentationPeiying Zhang, Nanxuan Zhao, Jing LiaoSIGGRAPH 2024 · 被引用 15 次
- Joint Implicit Neural Representation for High-fidelity and Compact Vector FontsChia-Hao Chen, Ying-Tian Liu, Zhifei Zhang, Yuan-Chen Guo 等ICCV 2023 · 被引用 4 次
