Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning
Dengming Zhang, Weitao You, Ziheng Liu, Lingyun Sun, Pei Chen
摘要
Dynamic Music Emotion Recognition (DMER) aims to predict the emotion of different moments in music, playing a crucial role in music information retrieval. The existing DMER methods struggle to capture long-term dependencies when dealing with sequence data, which limits their performance. Furthermore, these methods often overlook the influence of individual differences on emotion perception, even though everyone has their own personalized emotional perception in the real world. Motivated by these issues, we explore more effective sequence processing methods and introduce the Personalized DMER (PDMER) problem, which requires models to predict emotions that align with personalized perception. Specifically, we propose a Dual-Scale Attention-Based Meta-Learning (DSAML) method. This method fuses features from a dual-scale feature extractor and captures both short and long-term dependencies using a dual-scale attention transformer, improving the performance in traditional DMER. To achieve PDMER, we design a novel task construction strategy that divides tasks by annotators. Samples in a task are annotated by the same annotator, ensuring consistent perception. Leveraging this strategy alongside meta-learning, DSAML can predict personalized perception of emotions with just one personalized annotation sample. Our objective and subjective experiments demonstrate that our method can achieve state-of-the-art performance in both traditional DMER and PDMER.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- FG-Midiformer: A Symbolic Music Understanding Model towards Fine-Grained Learning of Multi-AttributesHaonan Cheng, Junwei Zhang, Hengyan Huang, Long YeACM MM 2025 · 被引用 1 次
- PEIA: Personality and Emotion Integrated Attentive Model for Music Recommendation on Social Media PlatformsTiancheng Shen, Jia Jia, Yan Li, Yihui Ma 等AAAI 2020 · 被引用 58 次
- Emotion-Based End-to-End Matching Between Image and Music in Valence-Arousal SpaceSicheng Zhao, Yaxian Li, Xingxu Yao, Weizhi Nie 等ACM MM 2020 · 被引用 30 次
- EMOE: Modality-Specific Enhanced Dynamic Emotion ExpertsYiyang Fang, Wenke Huang, Guancheng Wan, Kehua Su 等CVPR 2025
- Multi-Scale Similarity Aggregation for Dynamic Metric LearningDingyi Zhang, Yingming Li, Zhongfei ZhangACM MM 2023 · 被引用 2 次
