What Do Deep Saliency Models Learn about Visual Attention?
Shi Chen, Ming Jiang, Qi Zhao
摘要
In recent years, deep saliency models have made significant progress in predicting human visual attention. However, the mechanisms behind their success remain largely unexplained due to the opaque nature of deep neural networks. In this paper, we present a novel analytic framework that sheds light on the implicit features learned by saliency models and provides principled interpretation and quantification of their contributions to saliency prediction. Our approach decomposes these implicit features into interpretable bases that are explicitly aligned with semantic attributes and reformulates saliency prediction as a weighted combination of probability maps connecting the bases and saliency. By applying our framework, we conduct extensive analyses from various perspectives, including the positive and negative weights of semantics, the impact of training data and architectural designs, the progressive influences of fine-tuning, and common failure patterns of state-of-the-art deep saliency models. Additionally, we demonstrate the effectiveness of our framework by exploring visual attention characteristics in various application scenarios, such as the atypical attention of people with autism spectrum disorder, attention to emotion-eliciting stimuli, and attention evolution over time. Our code is publicly available at https://github.com/szzexpoi/saliency_analysis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Attend to Anything: Foundation Model for Unified Human Attention ModelingWenzhuo Zhao, Ronghao Xian, Keren Fu, Qijun ZhaoICML 2026
- Explainable Saliency: Articulating Reasoning with Contextual PrioritizationNuo Chen, Ming Jiang, Qi ZhaoCVPR 2025
它引用的顶会 Paper2
相关 Paper
- DANCE: Enhancing saliency maps using decoysYang Young Lu, Wenbo Guo, Xinyu Xing, William Stafford NobleICML 2021 · 被引用 14 次
- Regularized Pairwise Relationship based Analytics for Structured DataZhaojing Luo, Shaofeng Cai, Yatong Wang, Beng Chin OoiSIGMOD 2023 · 被引用 12 次
- Saliency-Guided Image TranslationLai Jiang, Mai Xu, Xiaofei Wang, Leonid SigalCVPR 2021
- Mesh Saliency: An Independent Perceptual Measure or a Derivative of Image Saliency?Ran Song, Wei Zhang, Yitian Zhao, Yonghuai Liu 等CVPR 2021
- Interpreting Internal Activation Patterns in Deep Temporal Neural Networks by Finding PrototypesSohee Cho, Wonjoon Chang, Ginkyeng Lee, Jaesik ChoiKDD 2021 · 被引用 10 次
