MoSound: An Interactive Tool for Generative Sound Design in Motion Graphics
Jialin Huang, Prem Seetharaman, Timothy Richard Langlois, Li-Yi Wei, Rubaiat Habib Kazi, Yotam I. Gingold
Abstract
Volume:
(A) User Interface with multiple audio events (B) Mapping x to stereo panning (C) Mapping v to volume Figure 1: Generating effects sounds from motion graphics videos via MoSound. Given a motion graphics video, our interactive system extracts the key events and generates the corresponding sound effects. (A) The user interface of our system, which includes views of the automatically extracted visual events, their timing, and descriptions of the graphics and motions. From there, the user can choose how to map the visual events to the sound effects, preview the generated sound effects, and export the final audio. (More details are in Figure 3.) (B,C) Several keyframes of an example video, along with the corresponding guide sounds from the extracted motions, suggested prompts, and final effect sounds generated by MoSound. Please see the accompanying video for animation and sound effects. Water video © Alejandro Imondi.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bed5b308-3ac4-4ebb-8db6-aa5b790ea4d0Builds on23
- Simple and Controllable Music GenerationJade Copet, Felix Kreuk, Itai Gat, Tal Remez et al.NeurIPS 2023 · 843 citations
- AudioLDM: Text-to-Audio Generation with Latent Diffusion ModelsHaohe Liu, Zehua Chen, Yi Yuan, Xinhao Mei et al.ICML 2023 · 773 citations
- Erasing Concepts from Diffusion ModelsRohit Gandikota, Joanna Materzynska, Jaden Fiotto-Kaufman, David BauICCV 2023 · 536 citations
- DITTO: Diffusion Inference-Time T-Optimization for Music GenerationZachary Novack, Julian J. McAuley, Taylor Berg-Kirkpatrick, Nicholas J. BryanICML 2024 · 81 citations
- Supporting Accessible Data Visualization Through Audio Data NarrativesAlexa F. Siu, Gene S.-H. Kim, Sile O'Modhrain, Sean FollmerCHI 2022 · 63 citations
Related papers
- AutoSFX: Automatic Sound Effect Generation for VideosYujia Wang, Zhongxu Wang, Hua HuangACM MM 2024 · 2 citations
- Soundify: Matching Sound Effects to VideoDavid Chuan-En Lin, Anastasis Germanidis, Cristóbal Valenzuela, Yining Shi et al.UIST 2023 · 15 citations
- SoundStager: Interactive Design of Story-Driven GenAI Soundscapes for VideoSuhyeon Yoo, Adolfo Hernandez Santisteban, Prem Seetharaman, Justin Salamon et al.CHI 2026 · 2 citations
- Eventfulness for Interactive Video AlignmentJiatian Sun, Longxiulin Deng, Triantafyllos Afouras, Andrew Owens et al.SIGGRAPH 2023 · 5 citations
- KinemaFX: A Kinematic-Driven Interactive System for Particle Effect Exploration and CustomizationYifei Zhang, Linping Yuan, Yuheng Zhao, Jielin Feng et al.UIST 2025 · 2 citations
