Optimal visual search based on a model of target detectability in natural images
Shima Rashidi, Krista A. Ehinger, Andrew Turpin, Lars Kulik
摘要
To analyse visual systems, the concept of an ideal observer promises an optimal response for a given task. Bayesian ideal observers can provide optimal responses under uncertainty, if they are given the true distributions as input. In visual search tasks, prior studies have used signal to noise ratio (SNR) or psychophysics experiments to set the distributional parameters for simple targets on backgrounds with known patterns, however these methods do not easily translate to complex targets on natural scenes. Here, we develop a model of target detectability in natural images to estimate the parameters of target-present and target-absent distributions for a visual search task. We present a novel approach for approximating the foveated detectability of a known target in natural backgrounds based on biological aspects of human visual system. Our model considers both the uncertainty about target position and the visual system's variability due to its reduced performance in the periphery compared to the fovea. Our automated prediction algorithm uses trained logistic regression as a post processing phase of a pre-trained deep neural network. Eye tracking data from 12 observers detecting targets on natural image backgrounds are used as ground truth to tune foveation parameters and evaluate the model, using cross-validation. Finally, the model of target detectability is used in a Bayesian ideal observer model of visual search, and compared to human search performance. * The code to reproduce the results of the paper can be found at https://github.com/rashidis/bio_ based_detectability 34th Conference on Neural Information Processing Systems (NeurIPS 2020), Vancouver, Canada.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Instant Reality: Gaze-Contingent Perceptual Optimization for 3D Virtual Reality StreamingShaoyu Chen, Budmonde Duinkharjav, Xin Sun, Li-Yi Wei 等IEEE VR 2022 · 被引用 25 次
- DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural ImagesOzgur Kara, Harris Nisar, James M. RehgNeurIPS 2025 · 被引用 7 次
- Unifying Top-Down and Bottom-Up Scanpath Prediction Using TransformersZhibo Yang, Sounak Mondal, Seoyoung Ahn, Ruoyu Xue 等CVPR 2024
相关 Paper
- Amodal Segmentation through Out-of-Task and Out-of-Distribution Generalization with a Bayesian ModelYihong Sun, Adam Kortylewski, Alan L. YuilleCVPR 2022 · 被引用 26 次
- Peripheral Vision TransformerJuhong Min, Yucheng Zhao, Chong Luo, Minsu ChoNeurIPS 2022 · 被引用 49 次
- Transformer brain encoders explain human high-level visual responsesHossein Adeli, Minni Sun, Nikolaus KriegeskorteNeurIPS 2025 · 被引用 14 次
- Tracking Without Re-recognition in Humans and MachinesDrew Linsley, Girik Malik, Junkyung Kim, Lakshmi Narasimhan Govindarajan 等NeurIPS 2021 · 被引用 21 次
- Modeling Human Visual Search Performance on Realistic Webpages Using Analytical and Deep Learning MethodsArianna Yuan, Yang LiCHI 2020 · 被引用 20 次
