HyFI: Hyperbolic Feature Interpolation for Brain-Vision Alignment
Sangmin Jo, Wootaek Jeong, Da-Woon Heo, Yoohwan Hwang, Heung-Il Suk
Abstract
Recent progress in artificial intelligence has encouraged numerous attempts to understand and decode human visual system from brain signals. These prior works typically align neural activity independently with semantic and perceptual features extracted from images using pre-trained vision models. However, they fail to account for two key challenges: (1) the modality gap arising from the natural difference in the information level of representation between brain signals and images, and (2) the fact that semantic and perceptual features are highly entangled within neural activity. To address these issues, we utilize hyperbolic space, which is well-suited for considering differences in the amount of information and has the geometric property that geodesics between two points naturally bend toward the origin, where the representational capacity is lower. Leveraging these properties, we propose a novel framework, Hyperbolic Feature Interpolation (HyFI), which interpolates between semantic and perceptual visual features along hyperbolic geodesics. This enables both the fusion and compression of perceptual and semantic information, effectively reflecting the limited expressiveness of brain signals and the entangled nature of these features. As a result, it facilitates better alignment between brain and visual features. We demonstrate that HyFI achieves state-of-the-art performance in zero-shot brain-to-image retrieval, outperforming prior methods with Top-1 accuracy improvements of up to +17.3% on THINGS-EEG and +9.1% on THINGS-MEG.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 40b62a46-5f65-45eb-b518-a4e6a62b9d9eBuilds on15
- Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation LearningWeixin Liang, Yuhui Zhang, Yongchan Kwon, Serena Yeung et al.NeurIPS 2022 · 834 citations
- Reconstructing the Mind's Eye: fMRI-to-Image with Contrastive Learning and Diffusion PriorsPaul S. Scotti, Atmadeep Banerjee, Jimmie Goode, Stepan Shabalin et al.NeurIPS 2023 · 282 citations
- Hyperbolic Image-text RepresentationsKaran Desai, Maximilian Nickel, Tanmay Rajpurohit, Justin Johnson et al.ICML 2023 · 137 citations
- Decoding Natural Images from EEG for Object RecognitionYonghao Song, Bingchuan Liu, Xiang Li, Nanlin Shi et al.ICLR 2024 · 135 citations
- Open Vocabulary Electroencephalography-to-Text Decoding and Zero-Shot Sentiment ClassificationZhenhailong Wang, Heng JiAAAI 2022 · 122 citations
Related papers
- MB2C: Multimodal Bidirectional Cycle Consistency for Learning Robust Visual Neural RepresentationsYayun Wei, Lei Cao, Hao Li, Yilin DongACM MM 2024 · 21 citations
- Shrinking the Teacher: An Adaptive Teaching Paradigm for Asymmetric EEG-Vision AlignmentLukun Wu, Jie Li, Ziqi Ren, Kaifan Zhang et al.AAAI 2026
- Leveraging Visual Blur Perception Characteristics for EEG DecodingWenchao Liu, Hongwei Li, Zhouyang Xu, Lin Ma et al.AAAI 2026
- Linguistic Priors for Visual Decoupling: Towards Symmetric Vision-Brain AlignmentDongjun Liu, Weichen Dai, Jingsheng Qian, Honggang Liu et al.CVPR 2026
- Human-Aligned Image Models Improve Visual Decoding from the BrainNona Rajabi, Antônio H. Ribeiro, Miguel Vasco, Farzaneh Taleb et al.ICML 2025
