Multi-Proxy Wasserstein Classifier for Image Classification
Benlin Liu, Yongming Rao, Jiwen Lu, Jie Zhou, Cho-Jui Hsieh
摘要
Most widely-used convolutional neural networks (CNNs) end up with a global average pooling layer and a fully-connected layer. In this pipeline, a certain class is represented by one template vector preserved in the feature banks of fully-connected layer. Yet, a class may have multiple properties useful for recognition while the above formulation only captures one of them. Therefore, it is desired to represent a class by multiple proxies. However, directly adding multiple linear layers turns out to be a trivial solution as no improvement can be observed. To tackle this problem, we adopt optimal transport theory to calculate a non-uniform matching flow between the elements in the feature map of a sample and the proxies of a class in a closed way. By doing so, the models are enabled to achieve partial matching as both the feature maps and the proxy set can now focus on a subset of elements from the counterpart. Such formulation also enables us to embed the samples into the Wasserstein metric space, which has many advantages over the original Euclidean space. This formulation can be achieved by a lightweight iterative algorithm, which can be easily embedded into the automatic differentiation framework. Empirical studies are performed on two widely-used classification datasets, CIFAR, and ILSVRC2012, and the substantial improvements on these two benchmarks demonstrate the effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- PLOT: Prompt Learning with Optimal Transport for Vision-Language ModelsGuangyi Chen, Weiran Yao, Xiangchen Song, Xinyue Li 等ICLR 2023 · 被引用 22 次
- Order Constraints in Optimal TransportFabian Lim, Laura Wynter, Shiau Hong LimICML 2022 · 被引用 4 次
- Detecting Open World Objects via Partial Attribute AssignmentMuli Yang, Gabriel James Goenawan, Huaiyuan Qin, Kai Han 等CVPR 2025
- OST: Refining Text Knowledge with Optimal Spatio-Temporal Descriptor for General Video RecognitionTom Tongjia Chen, Hongshan Yu, Zhengeng Yang, Zechuan Li 等CVPR 2024
它引用的顶会 Paper2
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Synchronizing Probability Measures on Rotations via Optimal TransportTolga Birdal, Michael Arbel, Umut Simsekli, Leonidas J. GuibasCVPR 2020
相关 Paper
- Template based Graph Neural Network with Optimal Transport DistancesCédric Vincent-Cuaz, Rémi Flamary, Marco Corneli, Titouan Vayer 等NeurIPS 2022 · 被引用 35 次
- Dynamic Hierarchical Mimicking Towards Consistent Optimization ObjectivesDuo Li, Qifeng ChenCVPR 2020
- Dataset Distillation via the Wasserstein MetricHaoyang Liu, Yijiang Li, Tiancheng Xing, Peiran Wang 等ICCV 2025 · 被引用 39 次
- POT: Prototypical Optimal Transport for Weakly Supervised Semantic SegmentationJian Wang, Tianhong Dai, Bingfeng Zhang, Siyue Yu 等CVPR 2025
- The Curse of Conditions: Analyzing and Improving Optimal Transport for Conditional Flow-Based GenerationHo Kei Cheng, Alexander Gerhard SchwingICCV 2025 · 被引用 1 次
