Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent Visibility
Yidi Li, Jun Xiao, Zhengda Lu, Yiqun Wang, Haiyong Jiang
Abstract
A DSLR photo of a time clock, clear pointer." "A DSLR photo of a LV handbag." "A flying dragon, highly detailed." "A lamma." "A crab." "An airplane." "A DSLR photo of a football helmet." "A campling." "A DSLR photo of a stylish Air Jordan shoes." "A space shuttle." "A rotary telephone." "An expensive office chair." Figure 1. Examples of multiview vector graphics generated by our method conditioned on text prompts, the first two rows are sketch and the bottom row is iconography. Our method is capable of producing vector graphics with consistent views, well-preserved shape structures, and accurate occlusion relationships. The non-visible curves are rendered with lower opacity for visualization. Small images on the top right of the sketch results are the corresponding view rendering from the auxiliary 3DGS [17] branch as a reference for shape structures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ce9d1ab5-c799-49b9-b4bb-12a3976e0ab0Cited by top-tier papers3
- ViewCraft3D: High-fidelity and View-Consistent 3D Vector Graphics SynthesisChuang Wang, Haitao Zhou, Ling Luo, Qian YuNeurIPS 2025 · 5 citations
- 3DrawAgent: Teaching LLM to Draw in 3D with Early Contrastive ExperienceHongcan Xiao, Xinyue Xiao, Yilin Wang, Yue Zhang et al.CVPR 2026 · 1 citation
- AffIn-Space: Learning Affine-Invariant Representations for 3D Spatial Understanding with MLLMsZhenyu Lu, Liupeng Li, Jinpeng Wang, Haoqian Kang et al.ICML 2026
Builds on39
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
Related papers
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu et al.CVPR 2024
- Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion ModelsRonghuan Wu, Wanchao Su, Jing LiaoCVPR 2025
- TAPS3D: Text-Guided 3D Textured Shape Generation from Pseudo SupervisionJiacheng Wei, Hao Wang, Jiashi Feng, Guosheng Lin et al.CVPR 2023
- Breathing Life Into Sketches Using Text-to-Video PriorsRinon Gal, Yael Vinker, Yuval Alaluf, Amit Bermano et al.CVPR 2024
- Acc3D: Accelerating Single Image to 3D Diffusion Models via Edge Consistency Guided Score DistillationKendong Liu, Zhiyu Zhu, Hui Liu, Junhui HouCVPR 2025
