Spotting the Unseen: Reciprocal Consensus Network Guided by Visual Archetypes
Wenbo Hu, Hongjian Zhan, Xinchen Ma, Yue Lu, Ching Y. Suen
Abstract
Humans often require only a few visual archetypes to spot novel objects. Based on this observation, we present a strategy rooted in ``spotting the unseen" by establishing dense correspondences between potential query image regions and a visual archetype, and we propose the Consensus Network (CoNet). Our method leverages relational patterns intra and inter images via Auto-Correlation Representation (ACR) and Mutual-Correlation Representation (MCR). Within each image, the ACR module is capable of encoding both local self-similarity and global context simultaneously. Between the query and support images, the MCR module computes the cross-correlation across two image representations and introduces a reciprocal consistency constraint, which can incorporate to exclude outliers and enhance model robustness. To overcome the challenges of low-resource training data, particularly in one-shot learning scenarios, we incorporate an adaptive margin strategy to better handle diverse instances. The experimental results indicate the effectiveness of the proposed method across diverse domains such as object detection in natural scenes, and text spotting in both historical manuscripts and natural scenes, which demonstrates its sparkling generalization ability. Our code is available at: https://github.com/infinite-hwb/conet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- Dynamic Context Correspondence Network for Semantic AlignmentShuaiyi Huang, Qiuyue Wang, Songyang Zhang, Shipeng Yan et al.ICCV 2019 · 97 citations
- Expanding Low-Density Latent Regions for Open-Set Object DetectionJiaming Han, Yuqiang Ren, Jian Ding, Xingjia Pan et al.CVPR 2022 · 84 citations
- Task-Adaptive Negative Envision for Few-Shot Open-Set RecognitionShiyuan Huang, Jiawei Ma, Guangxing Han, Shih-Fu ChangCVPR 2022 · 39 citations
- Few-Shot Object Detection With Attention-RPN and Multi-Relation DetectorQi Fan, Wei Zhuo, Chi-Keung Tang, Yu-Wing TaiCVPR 2020
- Few-Shot Open-Set Recognition by Transformation ConsistencyMinki Jeong, Seokeon Choi, Changick KimCVPR 2021
Related papers
- CRNet: Cross-Reference Networks for Few-Shot SegmentationWeide Liu, Chi Zhang, Guosheng Lin, Fayao LiuCVPR 2020
- Correspondence Networks With Adaptive Neighbourhood ConsensusShuda Li, Kai Han, Theo W. Costain, Henry Howard-Jenkins et al.CVPR 2020
- Query Adaptive Few-Shot Object Detection with Heterogeneous Graph Convolutional NetworksGuangxing Han, Yicheng He, Shiyuan Huang, Jiawei Ma et al.ICCV 2021 · 134 citations
- ConsNet: Learning Consistency Graph for Zero-Shot Human-Object Interaction DetectionYe Liu, Junsong Yuan, Chang Wen ChenACM MM 2020 · 83 citations
- Learning to Contextually Aggregate Multi-Source Supervision for Sequence LabelingOuyu Lan, Xiao Huang, Bill Yuchen Lin, He Jiang et al.ACL 2020 · 33 citations
