Cross-View Representation Learning for Multi-View Logo Classification with Information Bottleneck
Jing Wang, Yuanjie Zheng, Jingqi Song, Sujuan Hou
Abstract
Multi-view logo classification is a challenging task due to the cross-view misalignment of logo image varies under different viewpoints, large intra-classes and small inter-classes variation of logo appearance. Cross-view data can represent objects from different views and thus provide complementary information for data analysis. However, most existing multi-view algorithms usually maximize the correlation between different views for consistency. Those methods ignore the interaction among different views and may cause semantic bias during the process of common feature learning. In this paper, we investigate the information bottleneck (IB) to the multi-view learning for extracting the different view common features of one category, named Dual-View Information Bottleneck representation (Dual-view IB). To the best of our knowledge, this is the first cross-view learning method for logo classification. Specifically, we maximize the mutual information between the representations of the two views to achieve the preservation of key features in the classification task, while eliminating the redundant information that is not shared between the two views. In addition, due to the unbalance of samples and limited computing resources, we further introduce a novel Pair Batch Data Augmentation (PB) algorithm for Dual-view IB model, which applies augmentations from a learned policy based on replicates instances of two samples within the same batch. Comprehensive experiments on three existing benchmark datasets, which demonstrate the effectiveness of the proposed method that outperforms the methods in the state of the art. The proposed method is expected to further the development of cross-view representation learning.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers3
- Safe Multi-View Deep ClassificationWei Liu, Yufei Chen, Xiaodong Yue, Changqing Zhang et al.AAAI 2023 · 27 citations
- Alleviate Anchor-Shift: Explore Blind Spots with Cross-View Reconstruction for Incomplete Multi-View ClusteringSuyuan Liu, Siwei Wang, Ke Liang, Junpu Zhang et al.NeurIPS 2024 · 13 citations
- Prompt-Guided Alignment with Information Bottleneck Makes Image Compression Also a RestorerXuelin Shen, Quan Liu, Jiayin Xu, Wenhan YangNeurIPS 2025
Related papers
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman et al.ICLR 2020 · 330 citations
- Partial Multi-View Multi-Label Classification via Semantic Invariance Learning and Prototype ModelingChengliang Liu, Gehui Xu, Jie Wen, Yabo Liu et al.ICML 2024 · 18 citations
- Multi-View Information-Bottleneck Representation LearningZhibin Wan, Changqing Zhang, Pengfei Zhu, Qinghua HuAAAI 2021 · 116 citations
- Farewell to Mutual Information: Variational Distillation for Cross-Modal Person Re-IdentificationXudong Tian, Zhizhong Zhang, Shaohui Lin, Yanyun Qu et al.CVPR 2021
- Disentangled Cross-Modal Representation Learning with Enhanced Mutual SupervisionLu Gao, Wenlan Chen, Daoyuan Wang, Fei Guo et al.NeurIPS 2025 · 5 citations
