Neural Photofit: Gaze-based Mental Image Reconstruction
Florian Strohm, Ekta Sood, Sven Mayer, Philipp Müller, Mihai Bâce, Andreas Bulling
摘要
We propose a novel method that leverages human fixations to visually decode the image a person has in mind into a photofit (facial composite). Our method combines three neural networks: An encoder, a scoring network, and a decoder. The encoder extracts image features and predicts a neural activation map for each face looked at by a human observer. A neural scoring network compares the human and neural attention and predicts a relevance score for each extracted image feature. Finally, image features are aggregated into a single feature vector as a linear combination of all features weighted by relevance which a decoder de-codes into the final photofit. We train the neural scoring network on a novel dataset containing gaze data of 19 participants looking at collages of synthetic faces. We show that our method significantly outperforms a mean baseline predictor and report on a human study that shows that we can decode photofits that are visually plausible and close to the observer’s mental image.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Transformer brain encoders explain human high-level visual responsesHossein Adeli, Minni Sun, Nikolaus KriegeskorteNeurIPS 2025 · 被引用 14 次
- Human-Aligned Image Models Improve Visual Decoding from the BrainNona Rajabi, Antônio H. Ribeiro, Miguel Vasco, Farzaneh Taleb 等ICML 2025
- Neural encoding with visual attentionMeenakshi Khosla, Gia H. Ngo, Keith Jamison, Amy Kuceyeski 等NeurIPS 2020 · 被引用 6 次
- EVOKE: Efficient and High-Fidelity EEG-to-Video Reconstruction via Decoupling Implicit Neural RepresentationHaodong Jing, Panqi Yang, Dongyao Jiang, Zhipeng Liu 等AAAI 2026 · 被引用 1 次
- Usable and Fast Interactive Mental Face ReconstructionFlorian Strohm, Mihai Bâce, Andreas BullingUIST 2023 · 被引用 5 次
