Learning Partial Correlation based Deep Visual Representation for Image Classification
Saimunur Rahman, Piotr Koniusz, Lei Wang, Luping Zhou, Peyman Moghadam, Changming Sun
摘要
Visual representation based on covariance matrix has demonstrates its efficacy for image classification by characterising the pairwise correlation of different channels in convolutional feature maps. However, pairwise correlation will become misleading once there is another channel correlating with both channels of interest, resulting in the "confounding" effect. For this case, "partial correlation" which removes the confounding effect shall be estimated instead. Nevertheless, reliably estimating partial correlation requires to solve a symmetric positive definite matrix optimisation, known as sparse inverse covariance estimation (SICE). How to incorporate this process into CNN remains an open issue. In this work, we formulate SICE as a novel structured layer of CNN. To ensure end-to-end trainability, we develop an iterative method to solve the above matrix optimisation during forward and backward propagation steps. Our work obtains a partial correlation based deep visual representation and mitigates the small sample problem often encountered by covariance matrix estimation in CNN. Computationally, our model can be effectively trained with GPU and works well with a large number of channels of advanced CNNs. Experiments show the efficacy and superior classification performance of our deep visual representation compared to covariance matrix based counterparts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Learning Spatial-context-aware Global Visual Feature Representation for Instance Image RetrievalZhongyan Zhang, Lei Wang, Luping Zhou, Piotr KoniuszICCV 2023 · 被引用 13 次
- Understanding Matrix Function Normalizations in Covariance Pooling through the Lens of Riemannian GeometryZiheng Chen, Yue Song, Xiaojun Wu, Gaowen Liu 等ICLR 2025 · 被引用 1 次
- Riemannian High-Order Pooling for Brain Foundation ModelsChen Hu, Ziheng Chen, Rui Wang, Yefeng Zheng 等ICLR 2026
- Robust Distillation via Untargeted and Targeted Intermediate Adversarial SamplesJunhao Dong, Piotr Koniusz, Junxi Chen, Z. Jane Wang 等CVPR 2024
- 3Mformer: Multi-order Multi-mode Transformer for Skeletal Action RecognitionLei Wang, Piotr KoniuszCVPR 2023
它引用的顶会 Paper7
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Spectral Feature Augmentation for Graph Contrastive Learning and BeyondYifei Zhang, Hao Zhu, Zixing Song, Piotr Koniusz 等AAAI 2023 · 被引用 131 次
- Kernelized Few-shot Object Detection with Efficient Integral AggregationShan Zhang, Lei Wang, Naila Murray, Piotr KoniuszCVPR 2022 · 被引用 69 次
- Why Approximate Matrix Square Root Outperforms Accurate SVD in Global Covariance Pooling?Yue Song, Nicu Sebe, Wei WangICCV 2021 · 被引用 39 次
相关 Paper
- Learning Structured Gaussians to Approximate Deep EnsemblesIvor J. A. Simpson, Sara Vicente, Neill D. F. CampbellCVPR 2022 · 被引用 8 次
- A Weakly Supervised Fine Label Classifier Enhanced by Coarse SupervisionFariborz Taherkhani, Hadi Kazemi, Ali Dabouei, Jeremy M. Dawson 等ICCV 2019 · 被引用 30 次
- Invertible Concept-based Explanations for CNN Models with Non-negative Concept Activation VectorsRuihan Zhang, Prashan Madumal, Tim Miller, Krista A. Ehinger 等AAAI 2021 · 被引用 140 次
- PICNN: A Pathway towards Interpretable Convolutional Neural NetworksWengang Guo, Jiayi Yang, Huilin Yin, Qijun Chen 等AAAI 2024 · 被引用 6 次
- Learning Factorized Weight Matrix for Joint FilteringXiangyu Xu, Yongrui Ma, Wenxiu SunICML 2020 · 被引用 11 次
