Hypergraph-guided Intra- and Inter-category Relation Modeling for Fine-grained Visual Recognition
Lu Chen, Qiangchang Wang, Zhaohui Li, Yilong Yin
摘要
Fine-grained Visual Recognition (FGVR) aims to distinguish objects within similar subcategories. Humans adeptly perform this challenging task by leveraging both intra-category distinctiveness and inter-category similarity. However, previous methods fail to combine these two complementary dimensions and mine the intrinsic relations among various semantic features. To address these limitations, we propose HI2R, a Hypergraph-guided Intra- and Inter-category Relation Modeling approach, which simultaneously extracts the intra-category structural information and inter-category relation for more precise reasoning. Specifically, we exploit a Hypergraph-guided Structure Learning (HSL) module, which employs hypergraphs to capture high-order structural relations, transcending traditional graph-based methods that are limited to pairwise linkages. This advancement allows the model to adapt to significant intra-category variations. Additionally, we propose an Inter-category Relation Perception (IRP) module to improve feature discrimination across categories by extracting and analyzing semantic relations among them. Our objective is to alleviate the robustness issue associated with exclusive reliance on intra-category discriminative features. Furthermore, a random semantic consistency (RSC) loss is introduced to direct the model's attention to commonly overlooked yet distinctive regions, indirectly enhancing the representation ability of both HSL and IRP modules. Both qualitative and quantitative results demonstrate the effectiveness and usefulness of HI2R.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Analyzing and Boosting the Power of Fine-Grained Visual Recognition for Multi-modal Large Language ModelsHulingxiao He, Geng Li, Zijun Geng, Jinglin Xu 等ICLR 2025
- SLADE: Shielding against Dual Exploits in Large Vision-Language ModelsMd. Zarif Hossain, Ahmed ImteajCVPR 2025
- Disentangled Hypergraph-Guided Mamba Scanning for Fine-Grained Visual RecognitionZhongwei Xiong, Hao Wang, Xiaoyan Yu, Lingling Li 等AAAI 2026
相关 Paper
- Adversarial Reconstruction Feedback for Robust Fine-Grained GeneralizationShijie Wang, Jian Shi, Haojie LiICCV 2025 · 被引用 2 次
- Graph-Based High-Order Relation Discovery for Fine-Grained RecognitionYifan Zhao, Ke Yan, Feiyue Huang, Jia LiCVPR 2021
- DVF: Advancing Robust and Accurate Fine-Grained Image Retrieval with Retrieval GuidelinesXin Jiang, Hao Tang, Rui Yan, Jinhui Tang 等ACM MM 2024 · 被引用 18 次
- SIM-Trans: Structure Information Modeling Transformer for Fine-grained Visual CategorizationHongbo Sun, Xiangteng He, Yuxin PengACM MM 2022 · 被引用 128 次
- Category-specific Semantic Coherency Learning for Fine-grained Image RecognitionShijie Wang, Zhihui Wang, Haojie Li, Wanli OuyangACM MM 2020 · 被引用 23 次
