Exploring Figure-Ground Assignment Mechanism in Perceptual Organization
Wei Zhai, Yang Cao, Jing Zhang, Zheng-Jun Zha
Abstract
Perceptual organization is a challenging visual task that aims to perceive and group the individual visual element so that it is easy to understand the meaning of the scene as a whole. Most recent methods building upon advanced Convolutional Neural Network (CNN) come from learning discriminative representation and modeling context hierarchically. However, when the visual appearance difference between foreground and background is obscure, the performance of existing methods degrades significantly due to the visual ambiguity in the discrimination process. In this paper, we argue that the figure-ground assignment mechanism, which conforms to human vision cognitive theory, can be explored to empower CNN to achieve a robust perceptual organization despite visual ambiguity. Specifically, we present a novel Figure-Ground-Aided (FGA) module to learn the configural statistics of the visual scene and leverage it for the reduction of visual ambiguity. Particularly, we demonstrate the benefit of using stronger supervisory signals by teaching (FGA) module to perceive configural cues, i.e., convexity and lower region, that human deem important for the perceptual organization. Furthermore, an Interactive Enhancement Module (IEM) is devised to leverage such configural priors to assist representation learning, thereby achieving robust perception organization with complex visual ambiguities. In addition, a well-founded visual segregation test is designed to validate the capability of the proposed FGA mechanism explicitly. Comprehensive evaluation results demonstrate our proposed FGA mechanism can effectively enhance the capability of perception organization on various baseline models. Nevertheless, the model augmented via our proposed FGA mechanism also outperforms state-of-the-art approaches on four challenging real-world applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- Weakly-Supervised Concealed Object Segmentation with SAM-based Pseudo Labeling and Multi-scale Feature GroupingChunming He, Kai Li, Yachao Zhang, Guoxia Xu et al.NeurIPS 2023 · 205 citations
- Strategic Preys Make Acute Predators: Enhancing Camouflaged Object Detectors by Generating Camouflaged ObjectsChunming He, Kai Li, Yachao Zhang, Yulun Zhang et al.ICLR 2024 · 123 citations
- Evaluation and Improvement of Interpretability for Self-Explainable Part-Prototype NetworksQihan Huang, Mengqi Xue, Wenqi Huang, Haofei Zhang et al.ICCV 2023 · 47 citations
- Spatial-Aware Token for Weakly Supervised Object LocalizationPingyu Wu, Wei Zhai, Yang Cao, Jiebo Luo et al.ICCV 2023 · 19 citations
- PAID: Pairwise Angular-Invariant Decomposition for Continual Test-Time AdaptationKunyu Wang, Xueyang Fu, Yuanfei Bao, Chengjie Ge et al.NeurIPS 2025 · 7 citations
Builds on11
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao et al.ICCV 2019 · 1,054 citations
- ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive BiasYufei Xu, Qiming Zhang, Jing Zhang, Dacheng TaoNeurIPS 2021 · 429 citations
- Uncertainty-Guided Transformer Reasoning for Camouflaged Object DetectionFan Yang, Qiang Zhai, Xin Li, Rui Huang et al.ICCV 2021 · 293 citations
- Disentangling neural mechanisms for perceptual groupingJunkyung Kim, Drew Linsley, Kalpit Thakkar, Thomas SerreICLR 2020 · 61 citations
- LambdaNetworks: Modeling long-range Interactions without AttentionIrwan BelloICLR 2021 · 48 citations
Related papers
- PQA: Perceptual Question AnsweringYonggang Qi, Kai Zhang, Aneeshan Sain, Yi-Zhe SongCVPR 2021
- Deep Grouping Model for Unified Perceptual ParsingZhiheng Li, Wenxuan Bao, Jiayang Zheng, Chenliang XuCVPR 2020
- Psychologically-inspired, unsupervised inference of perceptual groups of GUI widgets from GUI imagesMulong Xie, Zhenchang Xing, Sidong Feng, Xiwei Xu et al.FSE 2022 · 30 citations
- Compositional Few-Shot Recognition with Primitive Discovery and EnhancingYixiong Zou, Shanghang Zhang, Ke Chen, Yonghong Tian et al.ACM MM 2020 · 30 citations
- Boat in the Sky: Background Decoupling and Object-aware Pooling for Weakly Supervised Semantic SegmentationJianjun Xu, Hongtao Xie, Hai Xu, Yuxin Wang et al.ACM MM 2022 · 13 citations
