Bridging Inter-Class Ambiguity and Spatial Variability in Flexible Object Recognition via Graph Distillation
Lin Zuo, Kunshan Yang, Mengmeng Jing, Xiangxu Zhao, Jiaqiao Chen
Abstract
Flexible object recognition remains challenging in multimedia scenarios due to inherently diverse shapes and sizes, and subtle inter-class differences. Graph-based vision models show promise in flexible objects recognition by capturing variable relationships. However, they suffer from two problems: (1) inter-class ambiguity hinders model discrimination and (2) frequent scale changes degrade model generalization. To address these limitations, we propose a unified graph distillation framework that enhances inter-class discrimination and spatial generalization while maintaining computational efficiency. For inter-class ambiguity problem, we introduce a virtual prototype module that dynamically generates learnable class prototypes via clustering intermediate features. These prototypes are incorporated into the distillation loss to sharpen decision boundaries. A global-local distillation mechanism further capture both image-level global semantics and patch-level local details, enhancing inter-class discrimination. For frequent scale changes problem, we design a patch-aware distillation strategy that transfers knowledge across multiple patch scales, strengthening the student model's spatial generalization to match various shapes and sizes of flexible objects, thus alleviate generalization degradation. Extensive experiments on flexible-object datasets (FDA, FSCW, CCSN) and challenging benchmarks (CIFAR-100, Mini-ImageNet) confirm effectiveness and efficiency of our method.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Related papers
- High-dimension Prototype is a Better Incremental Object Detection LearnerYanjie Wang, Liqun Chen, Tianming Zhao, Tao Zhang et al.ICLR 2025
- Scale-Equivalent Distillation for Semi-Supervised Object DetectionQiushan Guo, Yao Mu, Jianyu Chen, Tianqi Wang et al.CVPR 2022 · 40 citations
- Class Incremental Medical Image Segmentation via Prototype-Guided Calibration and Dual-Aligned DistillationShengqian Zhu, Chengrong Yu, Qiang Wang, Ying Song et al.AAAI 2026 · 1 citation
- SSA-Seg: Semantic and Spatial Adaptive Pixel-level Classifier for Semantic SegmentationXiaowen Ma, Zhenliang Ni, Xinghao ChenNeurIPS 2024
- Generalizable Knowledge Distillation from Vision Foundation Models for Semantic SegmentationChonghua Lv, Dong Zhao, Shuang Wang, Dou Quan et al.CVPR 2026 · 1 citation
