BrainRAM: Cross-Modality Retrieval-Augmented Image Reconstruction from Human Brain Activity
Dian Xie, Peiang Zhao, Jiarui Zhang, Kangqi Wei, Xiaobao Ni, Jiong Xia
摘要
Reconstructing visual stimuli from brain activities is crucial for deciphering the underlying mechanism of the human visual system. While recent studies have achieved notable results by leveraging deep generative models, challenges persist due to the lack of large-scale datasets and the inherent noise from non-invasive measurement methods. In this study, we draw inspiration from the mechanism of human memory and propose BrainRAM, a novel two-stage dual-guided framework for visual stimuli reconstruction. BrainRAM incorporates a Retrieval-Augmented Module (RAM) and diffusion prior to enhance the quality of reconstructed images from the brain. Specifically, in stage I, we transform fMRI voxels into the latent space of image and text embeddings via diffusion priors, obtaining preliminary estimates of the visual stimuli's semantics and structure. In stage II, based on previous estimates, we retrieve data from the LAION-2B-en dataset and employ the proposed RAM to refine them, yielding high-quality reconstruction results. Extensive experiments demonstrate that our BrainRAM outperforms current state-of-the-art methods both qualitatively and quantitatively, providing a new perspective for visual stimuli reconstruction.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Brain Image Reconstruction with Retrieval-Augmented DiffusionShuqi Zhu, Ziyi Ye, Yi Zhong, Qingyao Ai 等SIGIR 2025 · 被引用 1 次
- Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video ReconstructionYujie Wei, Chenglong Ma, Jianxiong Gao, Chenhui Wang 等CVPR 2026
相关 Paper
- Seeing Through the Brain: New Insights from Decoding Visual Stimuli with fMRIZheng Huang, Enpei Zhang, Weikang Qiu, Yinghao Cai 等ICLR 2026 · 被引用 2 次
- Mind Reader: Reconstructing complex images from brain activitiesSikun Lin, Thomas Sprague, Ambuj K. SinghNeurIPS 2022 · 被引用 155 次
- MindDiffuser: Controlled Image Reconstruction from Human Brain Activity with Semantic and Structural DiffusionYizhuo Lu, Changde Du, Qiongyi Zhou, Dianpeng Wang 等ACM MM 2023 · 被引用 48 次
- Contrast, Attend and Diffuse to Decode High-Resolution Images from Brain ActivitiesJingyuan Sun, Mingxiao Li, Zijiao Chen, Yunhao Zhang 等NeurIPS 2023 · 被引用 57 次
- Reconstructing the Mind's Eye: fMRI-to-Image with Contrastive Learning and Diffusion PriorsPaul S. Scotti, Atmadeep Banerjee, Jimmie Goode, Stepan Shabalin 等NeurIPS 2023 · 被引用 282 次
