Make Me Happier: Evoking Emotions through Image Diffusion Models
Qing Lin, Jingfeng Zhang, Yew-Soon Ong, Mengmi Zhang
Abstract
Despite the rapid progress in image generation, emotional image editing remains under-explored. The semantics, context, and structure of an image can evoke emotional responses, making emotional image editing techniques valuable for various real-world applications, including treatment of psychological disorders, commercialization of products, and artistic design. First, we present a novel challenge of emotion-evoked image generation, aiming to synthesize images that evoke target emotions while retaining the semantics and structures of the original scenes. To address this challenge, we propose a diffusion model capable of effectively understanding and editing source images to convey desired emotions and sentiments. Moreover, due to the lack of emotion editing datasets, we provide a unique dataset consisting of 340,000 pairs of images and their emotion annotations. Furthermore, we conduct human psychophysics experiments and introduce a new evaluation metric to systematically benchmark all the methods. Experimental results demonstrate that our method surpasses all competitive baselines. Our diffusion model is capable of identifying emotional cues from original images, editing images that elicit desired emotions, and meanwhile, preserving the semantic structure of the original images. All code, model, and dataset are available at GitHub.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f6683a4b-71af-4de4-aff3-eebb7aa0de29Cited by top-tier papers2
- Reimagining Personal Data: Unlocking the Potential of AI-Generated Images in Personal Data Meaning-MakingSoobin Park, Hankyung Kim, Youn-kyung LimCHI 2025 · 15 citations
- EmoStyle: Emotion-Driven Image StylizationJingyuan Yang, Zihuan Bai, Hui HuangCVPR 2026 · 2 citations
Builds on17
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song et al.ICLR 2022 · 2,128 citations
Related papers
- EmoEdit: Evoking Emotions through Image ManipulationJingyuan Yang, Jiawei Feng, Weibin Luo, Dani Lischinski et al.CVPR 2025
- EmIT: Emotional Interaction control in Text-to-image diffusion modelsHaofan Zhang, Shangfei WangACM MM 2025
- CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content GenerationKaishen Yuan, Yuting Zhang, Shang Gao, Yijie Zhu et al.ICLR 2026 · 10 citations
- Affective Image Filter: Reflecting Emotions from Text to ImagesShuchen Weng, Peixuan Zhang, Zheng Chang, Xinlong Wang et al.ICCV 2023 · 25 citations
- SAKR-Edit: Scene-Aware Knowledge Reasoning for Text-to-Image EditingJiawen Wang, Jianjun Li, Zhiyuan Ma, Ruixia BaiACM MM 2025
