Optimal visual search based on a model of target detectability in natural images
Shima Rashidi, Krista A. Ehinger, Andrew Turpin, Lars Kulik
Abstract
To analyse visual systems, the concept of an ideal observer promises an optimal response for a given task. Bayesian ideal observers can provide optimal responses under uncertainty, if they are given the true distributions as input. In visual search tasks, prior studies have used signal to noise ratio (SNR) or psychophysics experiments to set the distributional parameters for simple targets on backgrounds with known patterns, however these methods do not easily translate to complex targets on natural scenes. Here, we develop a model of target detectability in natural images to estimate the parameters of target-present and target-absent distributions for a visual search task. We present a novel approach for approximating the foveated detectability of a known target in natural backgrounds based on biological aspects of human visual system. Our model considers both the uncertainty about target position and the visual system's variability due to its reduced performance in the periphery compared to the fovea. Our automated prediction algorithm uses trained logistic regression as a post processing phase of a pre-trained deep neural network. Eye tracking data from 12 observers detecting targets on natural image backgrounds are used as ground truth to tune foveation parameters and evaluate the model, using cross-validation. Finally, the model of target detectability is used in a Bayesian ideal observer model of visual search, and compared to human search performance. * The code to reproduce the results of the paper can be found at https://github.com/rashidis/bio_ based_detectability 34th Conference on Neural Information Processing Systems (NeurIPS 2020), Vancouver, Canada.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Instant Reality: Gaze-Contingent Perceptual Optimization for 3D Virtual Reality StreamingShaoyu Chen, Budmonde Duinkharjav, Xin Sun, Li-Yi Wei et al.IEEE VR 2022 · 25 citations
- DiffEye: Diffusion-Based Continuous Eye-Tracking Data Generation Conditioned on Natural ImagesOzgur Kara, Harris Nisar, James M. RehgNeurIPS 2025 · 7 citations
- Unifying Top-Down and Bottom-Up Scanpath Prediction Using TransformersZhibo Yang, Sounak Mondal, Seoyoung Ahn, Ruoyu Xue et al.CVPR 2024
Related papers
- Amodal Segmentation through Out-of-Task and Out-of-Distribution Generalization with a Bayesian ModelYihong Sun, Adam Kortylewski, Alan L. YuilleCVPR 2022 · 26 citations
- Peripheral Vision TransformerJuhong Min, Yucheng Zhao, Chong Luo, Minsu ChoNeurIPS 2022 · 49 citations
- Transformer brain encoders explain human high-level visual responsesHossein Adeli, Minni Sun, Nikolaus KriegeskorteNeurIPS 2025 · 14 citations
- Tracking Without Re-recognition in Humans and MachinesDrew Linsley, Girik Malik, Junkyung Kim, Lakshmi Narasimhan Govindarajan et al.NeurIPS 2021 · 21 citations
- Modeling Human Visual Search Performance on Realistic Webpages Using Analytical and Deep Learning MethodsArianna Yuan, Yang LiCHI 2020 · 20 citations
