Deep-Modal: Real-Time Impact Sound Synthesis for Arbitrary Shapes
Xutong Jin, Sheng Li, Tianshu Qu, Dinesh Manocha, Guoping Wang
Abstract
Model sound synthesis is a physically-based sound synthesis method used to generate audio content in games and virtual worlds. We present a novel learning-based impact sound synthesis algorithm called Deep-Modal. Our approach can handle sound synthesis for common arbitrary objects, especially dynamic generated objects, in real-time. We present a new compact strategy to represent the mode data, corresponding to frequency and amplitude, as fixed-length vectors. This is combined with a new network architecture that can convert shape features of 3D objects into mode data. Our network is based on an encoder-decoder architecture with the contact positions of objects and external forces embedded. Our method can synthesize interactive sounds related to objects of various shapes at any contact position, as well as objects of different materials and sizes. The synthesis process only takes 0.01s on a GTX 1080 Ti GPU. We show the effectiveness of Deep-Modal through extensive evaluation using different metrics, including recall and precision of prediction, sound spectrogram, and a user study.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 493a6773-1b54-4d6e-aad8-df4facd3ad40Cited by top-tier papers8
- ObjectFolder 2.0: A Multisensory Object Dataset for Sim2Real TransferRuohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke et al.CVPR 2022 · 58 citations
- MESH2IR: Neural Acoustic Impulse Response Generator for Complex 3D ScenesAnton Ratnarajah, Zhenyu Tang, Rohith Aralikatti, Dinesh ManochaACM MM 2022 · 35 citations
- GWA: A Large High-Quality Acoustic Dataset for Audio ProcessingZhenyu Tang, Rohith Aralikatti, Anton Jeran Ratnarajah, Dinesh ManochaSIGGRAPH 2022 · 23 citations
- SonifyAR: Context-Aware Sound Generation in Augmented RealityXia Su, Jon E. Froehlich, Eunyee Koh, Chang XiaoUIST 2024 · 12 citations
- NeuralSound: learning-based modal sound synthesis with acoustic transferXutong Jin, Sheng Li, Guoping Wang, Dinesh ManochaSIGGRAPH 2022 · 11 citations
Builds on1
Related papers
- Physics-Driven Diffusion Models for Impact Sound Synthesis from VideosKun Su, Kaizhi Qian, Eli Shlizerman, Antonio Torralba et al.CVPR 2023
- DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference TasksXutong Jin, Chenxi Xu, Ruohan Gao, Jiajun Wu et al.SIGGRAPH 2024 · 3 citations
- Learning Acoustic Scattering Fields for Dynamic Interactive Sound PropagationZhenyu Tang, Hsien-Yu Meng, Dinesh ManochaIEEE VR 2021 · 13 citations
- REALIMPACT: A Dataset of Impact Sound Fields for Real ObjectsSamuel Clarke, Ruohan Gao, Mason L. Wang, Mark Rau et al.CVPR 2023
- Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion SimulationJin Woo Lee, Jaehyun Park, Min Jun Choi, Kyogu LeeNeurIPS 2024 · 9 citations
