Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion Models
Ronghuan Wu, Wanchao Su, Jing Liao
摘要
A car with lights emitting from it, on a road."
"A salmon sushi with wooden chopsticks and a dish of soy sauce." "A dog wears a chef's hat."
"A sunflower in bloom with grass and a butterfly hovering above."
"A red apple with green leaves, a worm in a hole, and a juice box with a straw."
"A pig wearing a backpack and a cowboy hat is standing on a skateboard." "An astronaut is riding a horse." "A sandcastle with a bucket, shovel and two seagulls flying above."
that Chat2SVG outperforms existing methods in visual fidelity, path regularity, and semantic alignment. Additionally, our system enables intuitive editing through natural language instructions, making professional vector graphics creation accessible to all users. Our code is available at https://chat2svg.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- OmniSVG: A Unified Scalable Vector Graphics Generation ModelYiying Yang, Wei Cheng, Sijin Chen, Xianfang Zeng 等NeurIPS 2025 · 被引用 90 次
- GeoGramBench: Benchmarking the Geometric Program Reasoning in Modern LLMsShixian Luo, Zhu zezhou, Yu Yuan, Yuncheng Yang 等ICLR 2026 · 被引用 15 次
- VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector AnimationGuotao Liang, Zhangcheng Wang, Chuang Wang, Juncheng Hu 等ICML 2026 · 被引用 11 次
- LottieGPT: Tokenizing Vector Animation for Autoregressive GenerationJunhao Chen, Kejun Gao, Yuehan Cui, Mingze Sun 等CVPR 2026 · 被引用 10 次
- MeshLLM: Empowering Large Language Models to Progressively Understand and Generate 3D MeshShuangkang Fang, I-Chao Shen, Yufeng Wang, Yi-Hsuan Tsai 等ICCV 2025 · 被引用 8 次
它引用的顶会 Paper25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song 等ICLR 2022 · 被引用 2,128 次
- ImageReward: Learning and Evaluating Human Preferences for Text-to-Image GenerationJiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong 等NeurIPS 2023 · 被引用 1,310 次
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 被引用 463 次
相关 Paper
- Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent VisibilityYidi Li, Jun Xiao, Zhengda Lu, Yiqun Wang 等CVPR 2025
- SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG GenerationHanqi Chen, Zhongyin Zhao, Ye Chen, Zhujin Liang 等ACM MM 2025 · 被引用 3 次
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu 等CVPR 2024
- Exploring Sparse MoE in GANs for Text-conditioned Image SynthesisJiapeng Zhu, Ceyuan Yang, Kecheng Zheng, Yinghao Xu 等CVPR 2025
- Breathing Life Into Sketches Using Text-to-Video PriorsRinon Gal, Yael Vinker, Yuval Alaluf, Amit Bermano 等CVPR 2024
