An Erudite Fine-Grained Visual Classification Model
Dongliang Chang, Yujun Tong, Ruoyi Du, Timothy M. Hospedales, Yi-Zhe Song, Zhanyu Ma
摘要
Current fine-grained visual classification (FGVC) models are isolated. In practice, we first need to identify the coarse-grained label of an object, then select the corresponding FGVC model for recognition. This hinders the application of FGVC algorithms in real-life scenarios. In this paper, we propose an erudite FGVC model jointly trained by several different datasets 1 , which can efficiently and accurately predict an object's fine-grained label across the combined label space. We found through a pilot study that positive and negative transfers co-occur when different datasets are mixed for training, i.e., the knowledge from other datasets is not always useful. Therefore, we first propose a feature disentanglement module and a feature re-fusion module to reduce negative transfer and boost positive transfer between different datasets. In detail, we reduce negative transfer by decoupling the deep features through many dataset-specific feature extractors. Subsequently, these are channel-wise re-fused to facilitate positive transfer. Finally, we propose a meta-learning based dataset-agnostic spatial attention layer to take full advantage of the multi-dataset training data, given that localisation is dataset-agnostic between different datasets. Experimental results across 11 different mixed-datasets built on four different FGVC datasets demonstrate the effectiveness of the proposed method. Furthermore, the proposed method can be easily combined with existing FGVC methods to obtain state-of-the-art results. Our code is available at https://github.com/PRIS-CV/An-Erudite- FGVC-Model. * indicates the corresponding author. 1 In this paper, different datasets mean different fine-grained visual classification datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Multi-View Active Fine-Grained Visual RecognitionRuoyi Du, Wenqing Yu, Heqing Wang, Ting-En Lin 等ICCV 2023 · 被引用 14 次
- Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous LearnerHanwen Zhong, Jiaxin Chen, Yutong Zhang, Di Huang 等NeurIPS 2024 · 被引用 9 次
- Seeing as Experts Do: A Knowledge-Augmented Agent for Open-Set Fine-Grained Visual UnderstandingJunhan Chen, Zilu Zhou, Yujun Tong, Dongliang Chang 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper9
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang 等ICCV 2019 · 被引用 2,239 次
- Learning Attentive Pairwise Interaction for Fine-Grained ClassificationPeiqin Zhuang, Yali Wang, Yu QiaoAAAI 2020 · 被引用 392 次
- Fine-Grained Recognition: Accounting for Subtle Differences between Similar ClassesGuolei Sun, Hisham Cholakkal, Salman H. Khan, Fahad Shahbaz Khan 等AAAI 2020 · 被引用 138 次
- Simple Multi-dataset DetectionXingyi Zhou, Vladlen Koltun, Philipp KrähenbühlCVPR 2022 · 被引用 85 次
相关 Paper
- Your "Flamingo" is My "Bird": Fine-Grained, or NotDongliang Chang, Kaiyue Pang, Yixiao Zheng, Zhanyu Ma 等CVPR 2021
- Data-free Knowledge Distillation for Fine-grained Visual CategorizationRenrong Shao, Wei Zhang, Jianhua Yin, Jun WangICCV 2023 · 被引用 7 次
- mDALU: Multi-Source Domain Adaptation and Label Unification with Partial DatasetsRui Gong, Dengxin Dai, Yuhua Chen, Wen Li 等ICCV 2021 · 被引用 27 次
- Filtration and Distillation: Enhancing Region Attention for Fine-Grained Visual CategorizationChuanbin Liu, Hongtao Xie, Zheng-Jun Zha, Lingfeng Ma 等AAAI 2020 · 被引用 179 次
- SIM-Trans: Structure Information Modeling Transformer for Fine-grained Visual CategorizationHongbo Sun, Xiangteng He, Yuxin PengACM MM 2022 · 被引用 128 次
