ContextSeg: Sketch Semantic Segmentation by Querying the Context with Attention
Jiawei Wang, Changjian Li
摘要
Sketch semantic segmentation is a well-explored and pivotal problem in computer vision involving the assignment of pre-defined part labels to individual strokes. This paper presents ContextSeg-a simple yet highly effective approach to tackling this problem with two stages. In the first stage, to better encode the shape and positional information of strokes, we propose to predict an extra dense distance field in an autoencoder network to reinforce structural information learning. In the second stage, we treat an entire stroke as a single entity and label a group of strokes within the same semantic part using an auto-regressive Transformer with the default attention mechanism. By group-based labeling, our method can fully leverage the context information when making decisions for the remaining groups of strokes. Our method achieves the best segmentation accuracy compared with state-of-the-art approaches on two representative datasets and has been extensively evaluated demonstrating its superior performance. Additionally, we offer insights into solving part imbalance in training data and the preliminary experiment on cross-category training, which can inspire future research in this field.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Instance Segmentation of Scene Sketches Using Natural Image PriorsMia Tang, Yael Vinker, Chuan Yan, Lvmin Zhang 等SIGGRAPH 2025 · 被引用 3 次
- VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch GenerationJiawei Wang, Zhiming Cui, Changjian LiICCV 2025 · 被引用 3 次
- SketchAgent: Language-Driven Sequential Sketch GenerationYael Vinker, Tamar Rott Shaham, Kristine Zheng, Alex Zhao 等CVPR 2025
它引用的顶会 Paper6
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Multiscale Vision TransformersHaoqi Fan, Bo Xiong, Karttikeya Mangalam, Yanghao Li 等ICCV 2021 · 被引用 1,611 次
- Free2CAD: parsing freehand drawings into CAD commandsChangjian Li, Hao Pan, Adrien Bousseau, Niloy J. MitraSIGGRAPH 2022 · 被引用 100 次
- Creative Sketch GenerationSongwei Ge, Vedanuj Goswami, Larry Zitnick, Devi ParikhICLR 2021
相关 Paper
- Open Vocabulary Semantic Scene Sketch UnderstandingAhmed Bourouis, Judith Ellen Fan, Yulia GryaditskayaCVPR 2024
- Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch GenerationRui Yang, Huining Li, Yiyi Long, Xiaojun Wu 等ICCV 2025 · 被引用 2 次
- CoSE: Compositional Stroke EmbeddingsEmre Aksan, Thomas Deselaers, Andrea Tagliasacchi, Otmar HilligesNeurIPS 2020 · 被引用 37 次
- Stroke Extraction of Chinese Character Based on Deep Structure Deformable Image RegistrationMeng Li, Yahan Yu, Yi Yang, Guanghao Ren 等AAAI 2023 · 被引用 7 次
- Generating Sketches in a Hierarchical Auto-Regressive Process for Flexible Sketch Drawing Manipulation at Stroke-LevelSicong Zang, Shuhui Gao, Zhijun FangAAAI 2026 · 被引用 2 次
