Deep Co-Attention Network for Multi-View Subspace Learning
Lecheng Zheng, Yu Cheng, Hongxia Yang, Nan Cao, Jingrui He
摘要
Many real-world applications involve data from multiple modalities and thus exhibit the view heterogeneity. For example, user modeling on social media might leverage both the topology of the underlying social network and the content of the users’ posts; in the medical domain, multiple views could be X-ray images taken at different poses. To date, various techniques have been proposed to achieve promising results, such as canonical correlation analysis based methods, etc. In the meanwhile, it is critical for decision-makers to be able to understand the prediction results from these methods. For example, given the diagnostic result that a model provided based on the X-ray images of a patient at different poses, the doctor needs to know why the model made such a prediction. However, state-of-the-art techniques usually suffer from the inability to utilize the complementary information of each view and to explain the predictions in an interpretable manner. To address these issues, in this paper, we propose a deep co-attention network for multi-view subspace learning, which aims to extract both the common information and the complementary information in an adversarial setting and provide robust interpretations behind the prediction to the end-users via the co-attention mechanism. In particular, it uses a novel cross reconstruction loss and leverages the label information to guide the construction of the latent representation by incorporating the classifier into our model. This improves the quality of latent representation and accelerates the convergence speed. Finally, we develop an efficient iterative algorithm to find the optimal encoders and discriminator, which are evaluated extensively on synthetic and real-world data sets. We also conduct a case study to demonstrate how the proposed method robustly interprets the predictions on an image data set.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Stacked Hybrid-Attention and Group Collaborative Learning for Unbiased Scene Graph GenerationXingning Dong, Tian Gan, Xuemeng Song, Jianlong Wu 等CVPR 2022 · 被引用 116 次
- MULAN: Multi-modal Causal Structure Learning and Root Cause Analysis for Microservice SystemsLecheng Zheng, Zhengzhang Chen, Jingrui He, Haifeng ChenWWW 2024 · 被引用 53 次
- Contrastive Learning with Complex HeterogeneityLecheng Zheng, Jinjun Xiong, Yada Zhu, Jingrui HeKDD 2022 · 被引用 29 次
- PURE: Positive-Unlabeled Recommendation with Generative Adversarial NetworkYao Zhou, Jianpeng Xu, Jun Wu, Zeinab Taghavi Nasrabadi 等KDD 2021 · 被引用 29 次
- Indirect Invisible Poisoning Attacks on Domain AdaptationJun Wu, Jingrui HeKDD 2021 · 被引用 15 次
它引用的顶会 Paper2
相关 Paper
- DC-SPAN: A Dual Contrastive Attention Network for Multi-View ClusteringJingyi Chen, Zhibin Dong, Tiejun Li, Yibo HanAAAI 2026
- Decomposition-Based Variational Network for Multi-Contrast MRI Super-Resolution and ReconstructionPengcheng Lei, Faming Fang, Guixu Zhang, Tieyong ZengICCV 2023 · 被引用 39 次
- Cross-View Mutual Learning for Semi-Supervised Medical Image SegmentationSong Wu, Xiaoyu Wei, Xinyue Chen, Yazhou Ren 等ACM MM 2024 · 被引用 16 次
- Deep Embedded Complementary and Interactive Information for Multi-View ClassificationJinglin Xu, Wenbin Li, Xinwang Liu, Dingwen Zhang 等AAAI 2020 · 被引用 65 次
- Dual Decomposition of Convex Optimization Layers for Consistent Attention in Medical ImagesTom Ron, Michal Weiler-Sagie, Tamir HazanICML 2022 · 被引用 7 次
