MusicBERT: A Self-supervised Learning of Music Representation
Hongyuan Zhu, Ye Niu, Di Fu, Hao Wang
摘要
Music recommendation has been one of the most used information retrieval services on internet. Finding suitable music for users' demands from tens of millions of music relies on the understanding of music content. Traditional studies usually focus on music representation based on massive user behavioral data and music meta-data, which ignore the audio characteristic of music. However, it is found that the melodic characteristics of music themselves can be further used to understand music. Moreover, how to utilize large-scale audio data to learn music representation is not well explored. To this end, we propose a self-supervised learning model for music representation. We firstly utilize a beat-level music pre-training model to learn the structure of music. Then, we use a multi-task learning framework to model music self-representation and co-relations between music, concurrently. Besides, we propose several downstream tasks to evaluate music representation, including music genre classification, music highlight, and music similarity retrieval. Extensive experiments on multiple music datasets demonstrate our model's superiority over baselines on learning music representation.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised TrainingYizhi Li, Ruibin Yuan, Ge Zhang, Yinghao Ma 等ICLR 2024 · 被引用 277 次
- A Domain-Knowledge-Inspired Music Embedding Space and a Novel Attention Mechanism for Symbolic Music ModelingZixun Guo, Jaeyong Kang, Dorien HerremansAAAI 2023 · 被引用 27 次
- Mimicking the Annotation Process for Recognizing the Micro ExpressionsBo-Kai Ruan, Ling Lo, Hong-Han Shuai, Wen-Huang ChengACM MM 2022 · 被引用 16 次
- MCSSME: Multi-Task Contrastive Learning for Semi-supervised Singing Melody Extraction from Polyphonic MusicShuai YuAAAI 2024 · 被引用 9 次
- DisCover: Disentangled Music Representation Learning for Cover Song IdentificationJiahao Xun, Shengyu Zhang, Yanting Yang, Jieming Zhu 等SIGIR 2023 · 被引用 7 次
相关 Paper
- MIDI-Zero: A MIDI-driven Self-Supervised Learning Approach for Music RetrievalYuhang Su, Wei Hu, Hongfeng Gao, Fan ZhangSIGIR 2025
- Exploiting Behavioral Consistence for Universal User RepresentationJie Gu, Feng Wang, Qinghui Sun, Zhiquan Ye 等AAAI 2021 · 被引用 32 次
- AudioMosaic: Contrastive Masked Audio Representation LearningHanxun Huang, Qizhou Wang, Xingjun Ma, Cihang Xie 等ICML 2026 · 被引用 2 次
- Contrastive Learning with Positive-Negative Frame Mask for Music RepresentationDong Yao, Zhou Zhao, Shengyu Zhang, Jieming Zhu 等WWW 2022 · 被引用 26 次
- It's Time for Artistic Correspondence in Music and VideoDídac Surís, Carl Vondrick, Bryan C. Russell, Justin SalamonCVPR 2022 · 被引用 33 次
