SVGformer: Representation Learning for Continuous Vector Graphics using Transformers
Defu Cao, Zhaowen Wang, Jose Echevarria, Yan Liu
Abstract
Advances in representation learning have led to great success in understanding and generating data in various domains. However, in modeling vector graphics data, the pure data-driven approach often yields unsatisfactory results in downstream tasks as existing deep learning methods often require the quantization of SVG parameters and cannot exploit the geometric properties explicitly. In this paper, we propose a transformer-based representation learning model (SVGformer) that directly operates on continuous input values and manipulates the geometric information of SVG to encode outline details and long-distance dependencies. SVGfomer can be used for various downstream tasks: reconstruction, classification, interpolation, retrieval, etc. We have conducted extensive experiments on vector font and icon datasets to show that our model can capture high-quality representation information and outperform the previous state-of-the-art on downstream tasks significantly.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39362ab8-26ae-4e5a-9ce3-10a803c12738Cited by top-tier papers10
- Rendering-Aware Reinforcement Learning for Vector Graphics GenerationJuan A. Rodríguez, Haotian Zhang, Abhay Puri, Rishav Pramanik et al.NeurIPS 2025 · 42 citations
- Text-to-Vector Generation with Neural Path RepresentationPeiying Zhang, Nanxuan Zhao, Jing LiaoSIGGRAPH 2024 · 15 citations
- SVGen: Interpretable Vector Graphics Generation with Large Language ModelsFeiyu Wang, Zhiyuan Zhao, Yuandong Liu, Da Zhang et al.ACM MM 2025 · 7 citations
- Bézier Splatting for Fast and Differentiable Vector Graphics RenderingXi Liu, Chaoyi Zhou, Nanxuan Zhao, Siyu HuangNeurIPS 2025 · 2 citations
- OmniLottie: Generating Vector Animations via Parameterized Lottie TokensYiying Yang, Wei Cheng, Sijin Chen, Honghao Fu et al.CVPR 2026 · 2 citations
Builds on18
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Spectral Temporal Graph Neural Network for Multivariate Time-series ForecastingDefu Cao, Yujing Wang, Juanyong Duan, Ce Zhang et al.NeurIPS 2020 · 841 citations
Related papers
- DeepSVG: A Hierarchical Generative Network for Vector Graphics AnimationAlexandre Carlier, Martin Danelljan, Alexandre Alahi, Radu TimofteNeurIPS 2020 · 247 citations
- Vector Grimoire: Codebook-based Shape Generation under Raster Image SupervisionMarco Cipriano, Moritz Feuerpfeil, Gerard de MeloICML 2025
- Sketchformer: Transformer-Based Representation for Sketched StructureLeo Sampaio Ferraz Ribeiro, Tu Bui, John P. Collomosse, Moacir PontiCVPR 2020
- Im2Vec: Synthesizing Vector Graphics Without Vector SupervisionPradyumna Reddy, Michaël Gharbi, Michal Lukác, Niloy J. MitraCVPR 2021
- Vector Calligrapher: Generating Scalable Vector Graphics via Structured Linguistic SupervisionBo Zhou, Xikang Chen, Yan Gong, Yin ZhangACL 2026
