Scene-Aware Background Music Synthesis
Yujia Wang, Wei Liang, Wanwan Li, Dingzeyu Li, Lap-Fai Yu
摘要
Background music not only provides auditory experience for users, but also conveys, guides, and promotes emotions that resonate with visual contents. Studies on how to synthesize background music for different scenes can promote research in many fields, such as human behaviour research. Although considerable effort has been directed toward music synthesis, the synthesis of appropriate music based on scene visual content remains an open problem.
In this paper we introduce an interactive background music synthesis algorithm guided by visual content. We leverage a cascading strategy to synthesize background music in two stages: Scene Visual Analysis and Background Music Synthesis. First, seeking a deep learning-based solution, we leverage neural networks to analyze the sentiment of the input scene. Second, real-time background music is synthesized by optimizing a cost function that guides the selection and transition of music clips to maximize the emotion consistency between visual and auditory criteria, and music continuity. In our experiments, we demonstrate the proposed approach can synthesize dynamic background music for different types of scenarios. We also conducted quantitative and qualitative analysis on the synthesized results of multiple example scenes to validate the efficacy of our approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Toward Automatic Audio Description Generation for Accessible VideosYujia Wang, Wei Liang, Haikun Huang, Yongqi Zhang 等CHI 2021 · 被引用 86 次
- SoundsRide: Affordance-Synchronized Music Mixing for In-Car Audio Augmented RealityMohamed Kari, Tobias Grosse-Puppendahl, Alexander Jagaciak, David Bethge 等UIST 2021 · 被引用 19 次
- RD-FGFS: A Rule-Data Hybrid Framework for Fine-Grained Footstep Sound Synthesis from Visual GuidanceQiutang Qi, Haonan Cheng, Yang Wang, Long Ye 等ACM MM 2023 · 被引用 3 次
- How Does it Sound?Kun Su, Xiulong Liu, Eli ShlizermanNeurIPS 2021 · 被引用 1 次
相关 Paper
- Mood-Driven Colorization of Virtual Indoor ScenesMichael Solah, Haikun Huang, Jiachuan Sheng, Tian Feng 等IEEE VR 2022 · 被引用 10 次
- MELFuSION: Synthesizing Music from Image and Language Cues Using Diffusion ModelsSanjoy Chowdhury, Sayan Nag, K. J. Joseph, Balaji Vasan Srinivasan 等CVPR 2024
- Dynamic Scene Adjustment Mechanism for Manipulating User Experience in VRYi Li, Zhitao Liu, Li Yuan, Haolan Tang 等IEEE VR 2024 · 被引用 3 次
- Climaxing VR Character with Scene-Aware Aesthetic Dress SynthesisSifan Hou, Yujia Wang, Bing Ning, Wei LiangIEEE VR 2021 · 被引用 5 次
- Novel-View Acoustic SynthesisChangan Chen, Alexander Richard, Roman Shapovalov, Vamsi Krishna Ithapu 等CVPR 2023
