Recognizing Vector Graphics without Rasterization
Xinyang Jiang, Lu Liu, Caihua Shan, Yifei Shen, Xuanyi Dong, Dongsheng Li
摘要
In this paper, we consider a different data format for images: vector graphics. In contrast to raster graphics which are widely used in image recognition, vector graphics can be scaled up or down into any resolution without aliasing or information loss, due to the analytic representation of the primitives in the document. Furthermore, vector graphics are able to give extra structural information on how low-level elements group together to form high level shapes or structures. These merits of graphic vectors have not been fully leveraged in existing methods. To explore this data format, we target on the fundamental recognition tasks: object localization and classification. We propose an efficient CNN-free pipeline that does not render the graphic into pixels (i.e. rasterization), and takes textual document of the vector graphics as input, called YOLaT (You Only Look at Text). YOLaT builds multi-graphs to model the structural and spatial information in vector graphics, and a dual-stream graph neural network is proposed to detect objects from the graph. Our experiments show that by directly operating on vector graphics, YOLaT outperforms raster-graphic based object detection baselines in terms of both average precision and efficiency. Code is available at https://github.com/microsoft/YOLaT- VectorGraphicsRecognition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Symbol as Points: Panoptic Symbol Spotting via Point-based RepresentationWenlong Liu, Tianyu Yang, Yuhan Wang, Qizhi Yu 等ICLR 2024 · 被引用 10 次
- Point or Line? Using Line-based Representation for Panoptic Symbol Spotting in CAD DrawingsXingguang Wei, Haomin Wang, Shenglong Ye, Ruifeng Luo 等NeurIPS 2025 · 被引用 5 次
- RendNet: Unified 2D/3D Recognizer with Latent Space RenderingRuoxi Shi, Xinyang Jiang, Caihua Shan, Yansen Wang 等CVPR 2022 · 被引用 4 次
- VGBench: Evaluating Large Language Models on Vector Graphics Understanding and GenerationBocheng Zou, Mu Cai, Jianrui Zhang, Yong Jae LeeEMNLP 2024 · 被引用 3 次
- Empowering LLMs to Understand and Generate Complex Vector GraphicsXiming Xing, Juncheng Hu, Guotao Liang, Jing Zhang 等CVPR 2025
它引用的顶会 Paper10
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 被引用 1,188 次
- DeepSVG: A Hierarchical Generative Network for Vector Graphics AnimationAlexandre Carlier, Martin Danelljan, Alexandre Alahi, Radu TimofteNeurIPS 2020 · 被引用 247 次
- A Learned Representation for Scalable Vector GraphicsRaphael Gontijo Lopes, David Ha, Douglas Eck, Jonathon ShlensICCV 2019 · 被引用 153 次
- Computer-Aided Design as LanguageYaroslav Ganin, Sergey Bartunov, Yujia Li, Ethan Keller 等NeurIPS 2021 · 被引用 129 次
相关 Paper
- CanvasVAE: Learning to Generate Vector Graphic DocumentsKota YamaguchiICCV 2021 · 被引用 103 次
- Im2Vec: Synthesizing Vector Graphics Without Vector SupervisionPradyumna Reddy, Michaël Gharbi, Michal Lukác, Niloy J. MitraCVPR 2021
- GAT-CADNet: Graph Attention Network for Panoptic Symbol Spotting in CAD DrawingsZhaohua Zheng, Jianfang Li, Lingjie Zhu, Honghua Li 等CVPR 2022 · 被引用 19 次
- VectorFloorSeg: Two-Stream Graph Attention Network for Vectorized Roughcast Floorplan SegmentationBingchen Yang, Haiyong Jiang, Hao Pan, Jun XiaoCVPR 2023
- SVGformer: Representation Learning for Continuous Vector Graphics using TransformersDefu Cao, Zhaowen Wang, Jose Echevarria, Yan LiuCVPR 2023
