Affect2MM: Affective Analysis of Multimedia Content Using Emotion Causality
Trisha Mittal, Puneet Mathur, Aniket Bera, Dinesh Manocha
Abstract
We present Affect2MM, a learning method for timeseries emotion prediction for multimedia content. Our goal is to automatically capture the varying emotions depicted by characters in real-life human-centric situations and behaviors. We use the ideas from emotion causation theories to computationally model and determine the emotional state evoked in clips of movies. Affect2MM explicitly models the temporal causality using attention-based methods and Granger causality. We use a variety of components like facial features of actors involved, scene understanding, visual aesthetics, action/situation description, and movie script to obtain an affective-rich representation to understand and perceive the scene. We use an LSTM-based learning model for emotion perception. To evaluate our method, we analyze and compare our performance on three datasets, SENDv1, MovieGraphs, and the LIRIS-ACCEDE dataset, and observe an average of 10 -15% increase in the performance over SOTA methods for all three datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d39e241e-c212-4966-8bf6-cfe1bd90642bCited by top-tier papers11
- Emotion-Prior Awareness Network for Emotional Video CaptioningPeipei Song, Dan Guo, Xun Yang, Shengeng Tang et al.ACM MM 2023 · 29 citations
- Temporal Sentiment Localization: Listen and Look in Untrimmed VideosZhicheng Zhang, Jufeng YangACM MM 2022 · 19 citations
- Dual-path Collaborative Generation Network for Emotional Video CaptioningCheng Ye, Weidong Chen, Jingyu Li, Lei Zhang et al.ACM MM 2024 · 15 citations
- ArtELingo: A Million Emotion Annotations of WikiArt with Emphasis on Diversity over Language and CultureYoussef Mohamed, Mohamed Abdelfattah, Shyma Alhuwaider, Feifan Li et al.EMNLP 2022 · 13 citations
- MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution DistillationZhicheng Zhang, Pancheng Zhao, Eunil Park, Jufeng YangCVPR 2024 · 11 citations
Builds on2
Related papers
- Enlarging the Long-time Dependencies via RL-based Memory Network in Movie Affective AnalysisJie Zhang, Yin Zhao, Kai QianACM MM 2022 · 4 citations
- Observe before Generate: Emotion-Cause aware Video Caption for Multimodal Emotion Cause Generation in ConversationsFanfan Wang, Heqing Ma, Xiangqing Shen, Jianfei Yu et al.ACM MM 2024 · 6 citations
- How You Feelin'? Learning Emotions and Mental States in Movie ScenesDhruv Srivastava, Aditya Kumar Singh, Makarand TapaswiCVPR 2023
- Emotion across Modalities and Cultures: Multilingual Multimodal Emotion-Cause Analysis with Memory-inspired FrameworkDan Wu, Xincheng Ju, Dong Zhang, Shoushan Li et al.ACM MM 2025 · 2 citations
- Pairwise Emotional Relationship Recognition in Drama Videos: Dataset and BenchmarkXun Gao, Yin Zhao, Jie Zhang, Longjun CaiACM MM 2021 · 8 citations
