Exploring Figure-Ground Assignment Mechanism in Perceptual Organization
Wei Zhai, Yang Cao, Jing Zhang, Zheng-Jun Zha
摘要
Perceptual organization is a challenging visual task that aims to perceive and group the individual visual element so that it is easy to understand the meaning of the scene as a whole. Most recent methods building upon advanced Convolutional Neural Network (CNN) come from learning discriminative representation and modeling context hierarchically. However, when the visual appearance difference between foreground and background is obscure, the performance of existing methods degrades significantly due to the visual ambiguity in the discrimination process. In this paper, we argue that the figure-ground assignment mechanism, which conforms to human vision cognitive theory, can be explored to empower CNN to achieve a robust perceptual organization despite visual ambiguity. Specifically, we present a novel Figure-Ground-Aided (FGA) module to learn the configural statistics of the visual scene and leverage it for the reduction of visual ambiguity. Particularly, we demonstrate the benefit of using stronger supervisory signals by teaching (FGA) module to perceive configural cues, i.e., convexity and lower region, that human deem important for the perceptual organization. Furthermore, an Interactive Enhancement Module (IEM) is devised to leverage such configural priors to assist representation learning, thereby achieving robust perception organization with complex visual ambiguities. In addition, a well-founded visual segregation test is designed to validate the capability of the proposed FGA mechanism explicitly. Comprehensive evaluation results demonstrate our proposed FGA mechanism can effectively enhance the capability of perception organization on various baseline models. Nevertheless, the model augmented via our proposed FGA mechanism also outperforms state-of-the-art approaches on four challenging real-world applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Weakly-Supervised Concealed Object Segmentation with SAM-based Pseudo Labeling and Multi-scale Feature GroupingChunming He, Kai Li, Yachao Zhang, Guoxia Xu 等NeurIPS 2023 · 被引用 205 次
- Strategic Preys Make Acute Predators: Enhancing Camouflaged Object Detectors by Generating Camouflaged ObjectsChunming He, Kai Li, Yachao Zhang, Yulun Zhang 等ICLR 2024 · 被引用 123 次
- Evaluation and Improvement of Interpretability for Self-Explainable Part-Prototype NetworksQihan Huang, Mengqi Xue, Wenqi Huang, Haofei Zhang 等ICCV 2023 · 被引用 47 次
- Spatial-Aware Token for Weakly Supervised Object LocalizationPingyu Wu, Wei Zhai, Yang Cao, Jiebo Luo 等ICCV 2023 · 被引用 19 次
- PAID: Pairwise Angular-Invariant Decomposition for Continual Test-Time AdaptationKunyu Wang, Xueyang Fu, Yuanfei Bao, Chengjie Ge 等NeurIPS 2025 · 被引用 7 次
它引用的顶会 Paper11
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao 等ICCV 2019 · 被引用 1,054 次
- ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive BiasYufei Xu, Qiming Zhang, Jing Zhang, Dacheng TaoNeurIPS 2021 · 被引用 429 次
- Uncertainty-Guided Transformer Reasoning for Camouflaged Object DetectionFan Yang, Qiang Zhai, Xin Li, Rui Huang 等ICCV 2021 · 被引用 293 次
- Disentangling neural mechanisms for perceptual groupingJunkyung Kim, Drew Linsley, Kalpit Thakkar, Thomas SerreICLR 2020 · 被引用 61 次
- LambdaNetworks: Modeling long-range Interactions without AttentionIrwan BelloICLR 2021 · 被引用 48 次
相关 Paper
- PQA: Perceptual Question AnsweringYonggang Qi, Kai Zhang, Aneeshan Sain, Yi-Zhe SongCVPR 2021
- Deep Grouping Model for Unified Perceptual ParsingZhiheng Li, Wenxuan Bao, Jiayang Zheng, Chenliang XuCVPR 2020
- Psychologically-inspired, unsupervised inference of perceptual groups of GUI widgets from GUI imagesMulong Xie, Zhenchang Xing, Sidong Feng, Xiwei Xu 等FSE 2022 · 被引用 30 次
- Compositional Few-Shot Recognition with Primitive Discovery and EnhancingYixiong Zou, Shanghang Zhang, Ke Chen, Yonghong Tian 等ACM MM 2020 · 被引用 30 次
- Boat in the Sky: Background Decoupling and Object-aware Pooling for Weakly Supervised Semantic SegmentationJianjun Xu, Hongtao Xie, Hai Xu, Yuxin Wang 等ACM MM 2022 · 被引用 13 次
