Zero-Shot Emotion Recognition via Affective Structural Embedding
Chi Zhan, Dongyu She, Sicheng Zhao, Ming-Ming Cheng, Jufeng Yang
摘要
Image emotion recognition attracts much attention in recent years due to its wide applications. It aims to understand the emotional response of humans, where candidate emotion categories are generally defined by specific psychological theories. However, with the development of psychological theories, emotion categories become increasingly diverse, fine-grained, and difficult to collect samples. In this paper, we investigate zero-shot learning (ZSL) problem in the emotion recognition task, which aims to recognize the new unseen emotions. Specifically, we propose an affective structural embedding framework, utilizing mid-level semantic representation, i.e., adjective-noun pairs (ANP) features, to construct an intermediate embedding space. By doing this, the learned intermediate space can bridge the affective gap between low-level visual features and high-level semantics. In addition, we introduce an adversarial constraint to combine the visual and affective embeddings so as to retain the discriminative capacity of visual features and the affective structural information of semantic features during training process. Our method is evaluated on five widelyused affective datasets and the experimental results show that the proposed algorithm outperforms the state-of-theart approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- An End-to-End Visual-Audio Attention Network for Emotion Recognition in User-Generated VideosSicheng Zhao, Yunsheng Ma, Yang Gu, Jufeng Yang 等AAAI 2020 · 被引用 123 次
- Emotion-Based End-to-End Matching Between Image and Music in Valence-Arousal SpaceSicheng Zhao, Yaxian Li, Xingxu Yao, Weizhi Nie 等ACM MM 2020 · 被引用 30 次
- Emotion-Prior Awareness Network for Emotional Video CaptioningPeipei Song, Dan Guo, Xun Yang, Shengeng Tang 等ACM MM 2023 · 被引用 29 次
- Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text PairsDaiqing Wu, Dongbao Yang, Yu Zhou, Can MaACM MM 2024 · 被引用 6 次
- Knowledge-Aligned Counterfactual-Enhancement Diffusion Perception for Unsupervised Cross-Domain Visual Emotion RecognitionWen Yin, Yong Wang, Guiduo Duan, Dongyang Zhang 等CVPR 2025
相关 Paper
- Consistent Structural Relation Learning for Zero-Shot SegmentationPeike Li, Yunchao Wei, Yi YangNeurIPS 2020 · 被引用 88 次
- Adaptive and Generative Zero-Shot LearningYu-Ying Chou, Hsuan-Tien Lin, Tyng-Luh LiuICLR 2021 · 被引用 25 次
- Generalized Zero-shot Learning with Multi-source Semantic Embeddings for Scene RecognitionXinhang Song, Haitao Zeng, Sixian Zhang, Luis Herranz 等ACM MM 2020 · 被引用 9 次
- Progressive Visual Content Understanding Network for Image Emotion ClassificationJicai Pan, Shangfei WangACM MM 2023 · 被引用 5 次
- A Variational Autoencoder with Deep Embedding Model for Generalized Zero-Shot LearningPeirong Ma, Xiao HuAAAI 2020 · 被引用 43 次
