RendNet: Unified 2D/3D Recognizer with Latent Space Rendering
Ruoxi Shi, Xinyang Jiang, Caihua Shan, Yansen Wang, Dongsheng Li
摘要
Vector graphics (VG) have been ubiquitous in our daily life with vast applications in engineering, architecture, designs, etc. The VG recognition process of most existing methods is to first render the VG into raster graphics (RG) and then conduct recognition based on RG formats. However, this procedure discards the structure of geometries and loses the high resolution of VG. Recently, another category of algorithms is proposed to recognize directly from the original VG format. But it is affected by the topological errors that can be filtered out by RG rendering. Instead of looking at one format, it is a good solution to utilize the formats of VG and RG together to avoid these shortcomings. Besides, we argue that the VG-to-RG rendering process is essential to effectively combine VG and RG information. By specifying the rules on how to transfer VG primitives to RG pixels, the rendering process depicts the interaction and correlation between VG and RG. As a result, we propose RendNet, a unified architecture for recognition on both 2D and 3D scenarios, which considers both VG/RG representations and exploits their interaction by incorporating the VG-to-RG rasterization process. Experiments show that Rend-Net can achieve state-of-the-art performance on 2D and 3D object recognition tasks on various VG datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Symbol as Points: Panoptic Symbol Spotting via Point-based RepresentationWenlong Liu, Tianyu Yang, Yuhan Wang, Qizhi Yu 等ICLR 2024 · 被引用 10 次
- VectorFloorSeg: Two-Stream Graph Attention Network for Vectorized Roughcast Floorplan SegmentationBingchen Yang, Haiyong Jiang, Hao Pan, Jun XiaoCVPR 2023
它引用的顶会 Paper12
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- DeepSVG: A Hierarchical Generative Network for Vector Graphics AnimationAlexandre Carlier, Martin Danelljan, Alexandre Alahi, Radu TimofteNeurIPS 2020 · 被引用 247 次
- A Learned Representation for Scalable Vector GraphicsRaphael Gontijo Lopes, David Ha, Douglas Eck, Jonathon ShlensICCV 2019 · 被引用 153 次
- Computer-Aided Design as LanguageYaroslav Ganin, Sergey Bartunov, Yujia Li, Ethan Keller 等NeurIPS 2021 · 被引用 129 次
- CoSE: Compositional Stroke EmbeddingsEmre Aksan, Thomas Deselaers, Andrea Tagliasacchi, Otmar HilligesNeurIPS 2020 · 被引用 37 次
相关 Paper
- Im2Vec: Synthesizing Vector Graphics Without Vector SupervisionPradyumna Reddy, Michaël Gharbi, Michal Lukác, Niloy J. MitraCVPR 2021
- Recognizing Vector Graphics without RasterizationXinyang Jiang, Lu Liu, Caihua Shan, Yifei Shen 等NeurIPS 2021 · 被引用 28 次
- UV-Net: Learning From Boundary RepresentationsPradeep Kumar Jayaraman, Aditya Sanghi, Joseph G. Lambourne, Karl D. D. Willis 等CVPR 2021
- VoGE: A Differentiable Volume Renderer using Gaussian Ellipsoids for Analysis-by-SynthesisAngtian Wang, Peng Wang, Jian Sun, Adam Kortylewski 等ICLR 2023 · 被引用 4 次
- VGBench: Evaluating Large Language Models on Vector Graphics Understanding and GenerationBocheng Zou, Mu Cai, Jianrui Zhang, Yong Jae LeeEMNLP 2024 · 被引用 3 次
