Where Can We Help? A Visual Analytics Approach to Diagnosing and Improving Semantic Segmentation of Movable Objects
Wenbin He, Lincan Zou, Arvind Kumar Shekar, Liang Gou, Liu Ren
Abstract
Semantic segmentation is a critical component in autonomous driving and has to be thoroughly evaluated due to safety concerns. Deep neural network (DNN) based semantic segmentation models are widely used in autonomous driving. However, it is challenging to evaluate DNN-based models due to their black-box-like nature, and it is even more difficult to assess model performance for crucial objects, such as lost cargos and pedestrians, in autonomous driving applications. In this work, we propose VASS, a Visual Analytics approach to diagnosing and improving the accuracy and robustness of Semantic Segmentation models, especially for critical objects moving in various driving scenes. The key component of our approach is a context-aware spatial representation learning that extracts important spatial information of objects, such as position, size, and aspect ratio, with respect to given scene contexts. Based on this spatial representation, we first use it to create visual summarization to analyze models' performance. We then use it to guide the generation of adversarial examples to evaluate models' spatial robustness and obtain actionable insights. We demonstrate the effectiveness of VASS via two case studies of lost cargo detection and pedestrian detection in autonomous driving. For both cases, we show quantitative evaluation on the improvement of models' performance with actionable insights obtained from VASS.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get df953fb0-8a83-4859-aa28-709141f084f1Cited by top-tier papers5
- DendroMap: Visual Exploration of Large-Scale Image Datasets for Machine Learning with TreemapsDonald Bertucci, Md Montaser Hamid, Yashwanthi Anand, Anita Ruangrotsakun et al.IEEE VIS 2022 · 34 citations
- Rapsai: Accelerating Machine Learning Prototyping of Multimedia Applications through Visual ProgrammingRuofei Du, Na Li, Jing Jin, Michelle Carney et al.CHI 2023 · 33 citations
- In Defence of Visual Analytics Systems: Replies to CriticsAoyu Wu, Dazhen Deng, Furui Cheng, Yingcai Wu et al.IEEE VIS 2022 · 30 citations
- A Unified Interactive Model Evaluation for Classification, Object Detection, and Instance Segmentation in Computer VisionChangjian Chen, Yukai Guo, Fengyuan Tian, Shilong Liu et al.IEEE VIS 2023 · 28 citations
- ModalChorus: Visual Probing and Alignment of Multi-Modal Embeddings via Modal Fusion MapYilin Ye, Shishi Xiao, Xingchen Zeng, Wei ZengIEEE VIS 2024 · 8 citations
Related papers
- VATLD: A Visual Analytics System to Assess, Understand and Improve Traffic Light DetectionLiang Gou, Lincan Zou, Nanxiang Li, Michael Hofmann et al.IEEE VIS 2020 · 72 citations
- Exploring Robustness of Unsupervised Domain Adaptation in Semantic SegmentationJinyu Yang, Chunyuan Li, Weizhi An, Hehuan Ma et al.ICCV 2021 · 34 citations
- Exploring the Adversarial Robustness of Video Object Segmentation via One-shot Adversarial AttacksKaixun Jiang, Lingyi Hong, Zhaoyu Chen, Pinxue Guo et al.ACM MM 2023 · 3 citations
- Semi-supervised Semantics-guided Adversarial Training for Robust Trajectory PredictionRuochen Jiao, Xiangguo Liu, Takami Sato, Qi Alfred Chen et al.ICCV 2023 · 26 citations
- DSRC: Learning Density-Insensitive and Semantic-Aware Collaborative Representation Against CorruptionsJingyu Zhang, Yilei Wang, Lang Qian, Peng Sun et al.AAAI 2025 · 13 citations
