Brain-Supervised Image Editing
Keith M. Davis, Carlos de la Torre-Ortiz, Tuukka Ruotsalo
摘要
Despite recent advances in deep neural models for semantic image editing, present approaches are dependent on explicit human input. Previous work assumes the availability of manually curated datasets for supervised learning, while for unsupervised approaches the human inspection of discovered components is required to identify those which modify worthwhile semantic features. Here, we present a novel alternative: the utilization of brain responses as a supervision signal for learning semantic feature representations. Participants <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> in a neurophysiological experiment were shown artificially generated faces and instructed to look for a particular semantic feature, such as “old” or “smiling”, while their brain responses were recorded via electroencephalography (EEG). Using supervision signals inferred from these responses, semantic features within the latent space of a generative adversarial network (GAN) were learned and then used to edit semantic features of new images. We show that implicit brain supervision achieves comparable semantic image editing performance to explicit manual labeling. This work demonstrates the feasibility of utilizing implicit human reactions recorded via brain-computer interfaces for semantic image editing and interpretation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Neural-Driven Image EditingPengfei Zhou, Jie Xia, Xiaopeng Peng, Wangbo Zhao 等NeurIPS 2025 · 被引用 5 次
- MindPilot: Closed-loop Visual Stimulation Optimization for Brain Modulation with EEG-guided DiffusionDongyang Li, Kunpeng Xie, Mingyang Wu, Yiwei Kong 等ICLR 2026
- When VR Meets BCI: (Un)Observable Brainwave-Aware Privacy Reconstruction in the Metaverse via Unrestricted Inbuilt Motion SensorsTao Ni, Zehua Sun, Qingchuan Zhao, Wei-Bin Lee 等S&P 2026
- MindPainter: Efficient Brain-Conditioned Painting of Natural Images via Cross-Modal Self-Supervised LearningMuzhou Yu, Shuyun Lin, Hongwei Yan, Kaisheng MaAAAI 2025
它引用的顶会 Paper15
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- GANalyze: Toward Visual Definitions of Cognitive Image PropertiesLore Goetschalckx, Alex Andonian, Aude Oliva, Phillip IsolaICCV 2019 · 被引用 345 次
- Controllable Artistic Text Style Transfer via Shape-Matching GANShuai Yang, Zhangyang Wang, Zhaowen Wang, Ning Xu 等ICCV 2019 · 被引用 110 次
- The Cortical Activity of Graded RelevanceZuzana Pinkosova, William J. McGeown, Yashar MoshfeghiSIGIR 2020 · 被引用 31 次
- Collaborative Filtering with Preferences Inferred from Brain SignalsKeith M. Davis III, Michiel M. A. Spapé, Tuukka RuotsaloWWW 2021 · 被引用 19 次
相关 Paper
- Cognition-Supervised Saliency Detection: Contrasting EEG Signals and Visual StimuliJun Ma, Tuukka RuotsaloACM MM 2024 · 被引用 2 次
- Brain Relevance Feedback for Interactive Image GenerationCarlos de la Torre-Ortiz, Michiel M. A. Spapé, Lauri Kangassalo, Tuukka RuotsaloUIST 2020 · 被引用 18 次
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
- Brainsourcing: Crowdsourcing Recognition Tasks via Collaborative Brain-Computer InterfacingKeith M. Davis, Lauri Kangassalo, Michiel M. A. Spapé, Tuukka RuotsaloCHI 2020 · 被引用 17 次
- Decoding Natural Images from EEG for Object RecognitionYonghao Song, Bingchuan Liu, Xiang Li, Nanlin Shi 等ICLR 2024 · 被引用 135 次
