Beyond Average: Individualized Visual Scanpath Prediction
Xianyu Chen, Ming Jiang, Qi Zhao
摘要
Understanding how attention varies across individuals has significant scientific and societal impacts. However, existing visual scanpath models treat attention uniformly, neglecting individual differences. To bridge this gap, this paper focuses on individualized scanpath prediction (ISP), a new attention modeling task that aims to accurately predict how different individuals shift their attention in diverse visual tasks. It proposes an ISP method featuring three novel technical components: (1) an observer encoder to characterize and integrate an observer's unique attention traits, (2) an observer-centric feature integration approach that holistically combines visual features, task guidance, and observer-specific characteristics, and (3) an adaptive fixation prioritization mechanism that refines scanpath predictions by dynamically prioritizing semantic feature maps based on individual observers' attention traits. These novel components allow scanpath models to effectively address the attention variations across different observers. Our method is generally applicable to different datasets, model architectures, and visual tasks, offering a comprehensive tool for transforming general scanpath models into individualized ones. Comprehensive evaluations using valuebased and ranking-based metrics verify the method's effectiveness and generalizability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural ImagesOzgur Kara, Harris Nisar, James M. RehgNeurIPS 2025 · 被引用 7 次
- What Moves the Eyes: Doubling Mechanistic Model Performance Using Deep Networks to Discover and Test Cognitive HypothesesFederico D'Agostino, Lisa Schwetlick, Matthias Bethge, Matthias KümmererNeurIPS 2025 · 被引用 4 次
- Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath PredictionGiuseppe Cartella, Vittorio Cuculo, Alessandro D'Amelio, Marcella Cornia 等ICCV 2025 · 被引用 3 次
- Gaze-Language Alignment for Zero-Shot Prediction of Visual Search Targets from Human Gaze ScanpathsSounak Mondal, Naveen Sendhilnathan, Ting Zhang, Yue Liu 等ICCV 2025 · 被引用 2 次
- Personalized Image Descriptions from Attention SequencesRuoyu Xue, Hieu Le, Jingyi Xu, Sounak Mondal 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper12
- UEyes: Understanding Visual Saliency across User Interface TypesYue Jiang, Luis A. Leiva, Hamed Rezazadegan Tavakoli, Paul R. B. Houssel 等CHI 2023 · 被引用 100 次
- Attention-Based Autism Spectrum Disorder Screening With Privileged ModalityShi Chen, Qi ZhaoICCV 2019 · 被引用 57 次
- Predicting Visual Importance Across Graphic Design TypesCamilo Fosco, Vincent Casser, Amish Kumar Bedi, Peter O'Donovan 等UIST 2020 · 被引用 55 次
- VisualHow: Multimodal Problem SolvingJinhui Yang, Xianyu Chen, Ming Jiang, Shi Chen 等CVPR 2022 · 被引用 7 次
- Learning the Best Pooling Strategy for Visual Semantic EmbeddingJiacheng Chen, Hexiang Hu, Hao Wu, Yuning Jiang 等CVPR 2021
相关 Paper
- Few-shot Personalized Scanpath PredictionRuoyu Xue, Jingyi Xu, Sounak Mondal, Hieu Le 等CVPR 2025
- Predicting Human Scanpaths in Visual Question AnsweringXianyu Chen, Ming Jiang, Qi ZhaoCVPR 2021
- Learning from Unique Perspectives: User-aware Saliency ModelingShi Chen, Nachiappan Valliappan, Shaolei Shen, Xinyu Ye 等CVPR 2023
- SpFormer: Spatio-Temporal Modeling for Scanpaths with TransformerWenqi Zhong, Linzhi Yu, Chen Xia, Junwei Han 等AAAI 2024 · 被引用 7 次
- Gazeformer: Scalable, Effective and Fast Prediction of Goal-Directed Human AttentionSounak Mondal, Zhibo Yang, Seoyoung Ahn, Dimitris Samaras 等CVPR 2023
