Rethinking Visual Reconstruction: Experience-Based Content Completion Guided by Visual Cues
Jiaxuan Chen, Yu Qi, Gang Pan
Abstract
Decoding seen images from brain activities has been an absorbing field. However, the reconstructed images still suffer from low quality with existing studies. This can be because our visual system is not like a camera that "remembers" every pixel. Instead, only part of the information can be perceived with our selective attention, and the brain "guesses" the rest to form what we think we see. Most existing approaches ignored the brain completion mechanism. In this work, we propose to reconstruct seen images with both the visual perception and the brain completion process, and design a simple, yet effective visual decoding framework to achieve this goal. Specifically, we first construct a shared discrete representation space for both brain signals and images. Then, a novel self-supervised token-to-token inpainting network is designed to implement visual content completion by building context and prior knowledge about the visual objects from the discrete latent space. Our approach improved the quality of visual reconstruction significantly and achieved state-of-the-art.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a9bb97fe-6b3a-476a-851f-90665d19e143Cited by top-tier papers6
- Bridging the Semantic Latent Space between Brain and Machine: Similarity Is All You NeedJiaxuan Chen, Yu Qi, Yueming Wang, Gang PanAAAI 2024 · 13 citations
- Beyond Brain Decoding: Visual-Semantic Reconstructions to Mental Creation Extension Based on fMRIHaodong Jing, Dongyao Jiang, Yongqiang Ma, Haibo Hua et al.ICCV 2025 · 6 citations
- Learning from Pattern Completion: Self-supervised Controllable GenerationZhiqiang Chen, Guofan Fan, Jinying Gao, Lei Ma et al.NeurIPS 2024 · 1 citation
- Bridging the Gap Between Brain and Machine in Interpreting Visual Semantics: Towards Self-Adaptive Brain-to-Text DecodingJiaxuan Chen, Yu Qi, Yueming Wang, Gang PanICCV 2025 · 1 citation
- Mind Artist: Creating Artistic Snapshots with Human ThoughtJiaxuan Chen, Yu Qi, Yueming Wang, Gang PanCVPR 2024
Builds on8
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Palette: Image-to-Image Diffusion ModelsChitwan Saharia, William Chan, Huiwen Chang, Chris A. Lee et al.SIGGRAPH 2022 · 1,638 citations
- Instance-Conditioned GANArantxa Casanova, Marlène Careil, Jakob Verbeek, Michal Drozdzal et al.NeurIPS 2021 · 167 citations
- Reconstructing Perceptive Images from Brain Activity by Shape-Semantic GANTao Fang, Yu Qi, Gang PanNeurIPS 2020 · 69 citations
Related papers
- Seeing Beyond the Brain: Conditional Diffusion Model with Sparse Masked Modeling for Vision DecodingZijiao Chen, Jiaxin Qing, Tiange Xiang, Wan Lin Yue et al.CVPR 2023
- Conditional Generative Neural Decoding with Structured CNN Feature PredictionChangde Du, Changying Du, Lijie Huang, Huiguang HeAAAI 2020 · 14 citations
- EVOKE: Efficient and High-Fidelity EEG-to-Video Reconstruction via Decoupling Implicit Neural RepresentationHaodong Jing, Panqi Yang, Dongyao Jiang, Zhipeng Liu et al.AAAI 2026 · 1 citation
- Mind Reader: Reconstructing complex images from brain activitiesSikun Lin, Thomas Sprague, Ambuj K. SinghNeurIPS 2022 · 155 citations
- Contrast, Attend and Diffuse to Decode High-Resolution Images from Brain ActivitiesJingyuan Sun, Mingxiao Li, Zijiao Chen, Yunhao Zhang et al.NeurIPS 2023 · 57 citations
