Farewell to Mutual Information: Variational Distillation for Cross-Modal Person Re-Identification
Xudong Tian, Zhizhong Zhang, Shaohui Lin, Yanyun Qu, Yuan Xie, Lizhuang Ma
Abstract
The Information Bottleneck (IB) provides an information theoretic principle for representation learning, by retaining all information relevant for predicting label while minimizing the redundancy. Though IB principle has been applied to a wide range of applications, its optimization remains a challenging problem which heavily relies on the accurate estimation of mutual information. In this paper, we present a new strategy, Variational Self-Distillation (VSD), which provides a scalable, flexible and analytic solution to essentially fitting the mutual information but without explicitly estimating it. Under rigorously theoretical guarantee, VSD enables the IB to grasp the intrinsic correlation between representation and label for supervised training. Furthermore, by extending VSD to multi-view learning, we introduce two other strategies, Variational Cross-Distillation (VCD) and Variational Mutual-Learning (VML), which significantly improve the robustness of representation to viewchanges by eliminating view-specific and task-irrelevant information. To verify our theoretically grounded strategies, we apply our approaches to cross-modal person Re-ID, and conduct extensive experiments, where the superior performance against state-of-the-art methods are demonstrated. Our intriguing findings highlight the need to rethink the way to estimate mutual information.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0f911bd3-08c5-4b78-909e-4faa4819bed7Cited by top-tier papers25
- Learning with Twin Noisy Labels for Visible-Infrared Person Re-IdentificationMouxing Yang, Zhenyu Huang, Peng Hu, Taihao Li et al.CVPR 2022 · 248 citations
- Learning Memory-Augmented Unidirectional Metrics for Cross-modality Person Re-identificationJialun Liu, Yifan Sun, Feng Zhu, Hongbin Pei et al.CVPR 2022 · 196 citations
- Exposing the Deception: Uncovering More Forgery Clues for Deepfake DetectionZhongjie Ba, Qingyu Liu, Zhenguang Liu, Shuang Wu et al.AAAI 2024 · 101 citations
- Cross-Domain Correlation Distillation for Unsupervised Domain Adaptation in Nighttime Semantic SegmentationHuan Gao, Jichang Guo, Guoli Wang, Qian ZhangCVPR 2022 · 82 citations
- Learning Modal-Invariant and Temporal-Memory for Video-based Visible-Infrared Person Re-IdentificationXinyu Lin, Jinxing Li, Zeyu Ma, Huafeng Li et al.CVPR 2022 · 81 citations
Builds on8
- Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self DistillationLinfeng Zhang, Jiebo Song, Anni Gao, Jingwei Chen et al.ICCV 2019 · 1,069 citations
- RGB-Infrared Cross-Modality Person Re-Identification via Joint Pixel and Feature AlignmentGuan'an Wang, Tianzhu Zhang, Jian Cheng, Si Liu et al.ICCV 2019 · 464 citations
- Infrared-Visible Cross-Modal Person Re-Identification with an X ModalityDiangang Li, Xing Wei, Xiaopeng Hong, Yihong GongAAAI 2020 · 419 citations
- Cross-Modality Paired-Images Generation for RGB-Infrared Person Re-IdentificationGuan'an Wang, Tianzhu Zhang, Yang Yang, Jian Cheng et al.AAAI 2020 · 364 citations
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman et al.ICLR 2020 · 330 citations
Related papers
- Connecting Jensen-Shannon and Kullback-Leibler Divergences: A New Bound for Representation LearningReuben Dorent, Polina Golland, William (Sandy) WellsNeurIPS 2025 · 7 citations
- Differentiable Information Bottleneck for Deterministic Multi-View ClusteringXiaoqiang Yan, Zhixiang Jin, Fengshou Han, Yangdong YeCVPR 2024 · 19 citations
- RedCore: Relative Advantage Aware Cross-Modal Representation Learning for Missing Modalities with Imbalanced Missing RatesJun Sun, Xinxin Zhang, Shoukang Han, Yu-Ping Ruan et al.AAAI 2024
- IBMA: Information Bottleneck-Based Multimodal AlignmentYancheng Wang, Zeyu Dong, Dongfang Sun, Alvin Silva et al.ICML 2026
- Permutation-Consistent Variational Encoding for Incomplete Multi-View Multi-Label ClassificationChengliang Liu, Bo Li, Bob Zhang, Xiaoling Luo et al.ICLR 2026
