PDiscoNet: Semantically consistent part discovery for fine-grained recognition
Robert van der Klis, Stephan Alaniz, Massimiliano Mancini, Cássio Fraga Dantas, Dino Ienco, Zeynep Akata, Diego Marcos
摘要
Fine-grained classification often requires recognizing specific object parts, such as beak shape and wing patterns for birds. Encouraging a fine-grained classification model to first detect such parts and then using them to infer the class could help us gauge whether the model is indeed looking at the right details better than with interpretability methods that provide a single attribution map. We propose PDiscoNet to discover object parts by using only image-level class labels along with priors encouraging the parts to be: discriminative, compact, distinct from each other, equivariant to rigid transforms, and active in at least some of the images. In addition to using the appropriate losses to encode these priors, we propose to use part-dropout, where full part feature vectors are dropped at once to prevent a single part from dominating in the classification, and part feature vector modulation, which makes the information coming from each part distinct from the perspective of the classifier. Our results on CUB, CelebA, and PartImageNet show that the proposed method provides substantially better part discovery performance than previous methods while not requiring any additional hyper-parameter tuning and without penalizing the classification performance. The code is available at https: //github.com/robertdvdk/part_detection
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- SUB: Benchmarking CBM Generalization via Synthetic Attribute SubstitutionsJessica Bader, Leander Girrbach, Stephan Alaniz, Zeynep AkataICCV 2025 · 被引用 8 次
- Low-Resource Vision Challenges for Foundation ModelsYunhua Zhang, Hazel Doughty, Cees G. M. SnoekCVPR 2024 · 被引用 7 次
- Understanding Multi-Granularity for Open-Vocabulary Part SegmentationJiho Choi, Seonho Lee, Seungho Lee, Minhyun Lee 等NeurIPS 2024 · 被引用 7 次
- PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part SegmentationJianjian Yin, Tao Chen, Yi Chen, Gensheng Pei 等CVPR 2026 · 被引用 6 次
- Unsupervised Part Discovery via Descriptor-Based Masked Image Restoration with Optimized ConstraintsJiahao Xia, Yike Wu, Wenjian Huang, Jianguo Zhang 等ICCV 2025 · 被引用 1 次
它引用的顶会 Paper7
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Part Discovery from Contrastive ReconstructionSubhabrata Choudhury, Iro Laina, Christian Rupprecht, Andrea VedaldiNeurIPS 2021 · 被引用 74 次
- B-cos Networks: Alignment is All We Need for InterpretabilityMoritz Böhle, Mario Fritz, Bernt SchieleCVPR 2022 · 被引用 62 次
- Part-Based Models Improve Adversarial RobustnessChawin Sitawarin, Kornrapat Pongmala, Yizheng Chen, Nicholas Carlini 等ICLR 2023 · 被引用 2 次
- Convolutional Dynamic Alignment Networks for Interpretable ClassificationsMoritz Böhle, Mario Fritz, Bernt SchieleCVPR 2021
相关 Paper
- Interpretable and Accurate Fine-grained Recognition via Region GroupingZixuan Huang, Yin LiCVPR 2020
- Unsupervised Part Segmentation Through Disentangling Appearance and ShapeShilong Liu, Lei Zhang, Xiao Yang, Hang Su 等CVPR 2021
- Post-hoc Part-Prototype NetworksAndong Tan, Fengtao Zhou, Hao ChenICML 2024 · 被引用 7 次
- PIP-Net: Patch-Based Intuitive Prototypes for Interpretable Image ClassificationMeike Nauta, Jörg Schlötterer, Maurice van Keulen, Christin SeifertCVPR 2023
- Dual Part Discovery Network for Zero-Shot LearningJiannan Ge, Hongtao Xie, Shaobo Min, Pandeng Li 等ACM MM 2022 · 被引用 21 次
