Bridging the Vision-Brain Gap with an Uncertainty-Aware Blur Prior
Haitao Wu, Qing Li, Changqing Zhang, Zhen He, Xiaomin Ying
Abstract
Can our brain signals faithfully reflect the original visual stimuli, even including high-frequency details? Although human perceptual and cognitive capacities enable us to process and remember visual information, these abilities are constrained by several factors, such as limited attentional resources and the finite capacity of visual memory. When visual stimuli are processed by human visual system into brain signals, some information is inevitably lost, leading to a discrepancy known as the System GAP. Additionally, perceptual and cognitive dynamics, along with technical noise in signal acquisition, degrade the fidelity of brain signals relative to the visual stimuli, known as the Random GAP. When encoded brain representations are directly aligned with the corresponding pretrained image features, the System GAP and Random GAP between paired data challenge the model, requiring it to bridge these gaps. However, in the context of limited paired data, these gaps are difficult for the model to learn, leading to overfitting and poor generalization to new data. To address these GAPs, we propose a simple yet effective approach called the Uncertainty-aware Blur Prior (UBP). It estimates the uncertainty within the paired data, reflecting the mismatch between brain signals and visual stimuli. Based on this uncertainty, UBP dynamically blurs the high-frequency details of the original images, reducing the impact of the mismatch and improving alignment. Our method achieves a top-1 accuracy of 50.9% and a top-5 accuracy of 79.7% on the zero-shot brain-to-image retrieval task, surpassing previous state-of-the-art methods by margins of 13.7% and 9.8%, respectively. Code is available at GitHub.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6d34193a-e5cf-4ce3-8f35-de4ee30dfa25Cited by top-tier papers10
- NeuroBridge: Bio-Inspired Self-Supervised EEG-to-Image Decoding via Cognitive Priors and Bidirectional Semantic AlignmentWenjiang Zhang, Sifeng Wang, Yuwei Su, Xinyu Li et al.AAAI 2026 · 7 citations
- EEGMirror: Leveraging EEG Data in the Wild Via Montage-Agnostic Self-Supervision for EEG to Video DecodingXuan-Hao Liu, Bao-Liang Lu, Wei-Long ZhengICCV 2025 · 5 citations
- Learning Brain Representation with Hierarchical Visual EmbeddingsJiawen Zheng, Haonan Jia, MING LI, Yuhui Zheng et al.ICLR 2026 · 3 citations
- HyFI: Hyperbolic Feature Interpolation for Brain-Vision AlignmentSangmin Jo, Wootaek Jeong, Da-Woon Heo, Yoohwan Hwang et al.AAAI 2026 · 2 citations
- D-FOSA: Dual-Diffusion Guided EEG-to-Image Reconstruction with Frequency-Oriented Semantic AlignmentChenglong Yu, Shuai Shen, Xiangsheng Li, Yang LiCVPR 2026 · 1 citation
Builds on26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen et al.ICML 2021 · 5,401 citations
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo et al.ICCV 2019 · 1,125 citations
Related papers
- Shrinking the Teacher: An Adaptive Teaching Paradigm for Asymmetric EEG-Vision AlignmentLukun Wu, Jie Li, Ziqi Ren, Kaifan Zhang et al.AAAI 2026
- Linguistic Priors for Visual Decoupling: Towards Symmetric Vision-Brain AlignmentDongjun Liu, Weichen Dai, Jingsheng Qian, Honggang Liu et al.CVPR 2026
- Leveraging Visual Blur Perception Characteristics for EEG DecodingWenchao Liu, Hongwei Li, Zhouyang Xu, Lin Ma et al.AAAI 2026
- Towards Brain Passage Retrieval: An Investigation of EEG Query RepresentationsNiall McGuire, Yashar MoshfeghiSIGIR 2025 · 5 citations
- Adaptive Uncertainty-Based Learning for Text-Based Person RetrievalShenshen Li, Chen He, Xing Xu, Fumin Shen et al.AAAI 2024 · 59 citations
