REALIMPACT: A Dataset of Impact Sound Fields for Real Objects
Samuel Clarke, Ruohan Gao, Mason L. Wang, Mark Rau, Julia Xu, Jui-Hsien Wang, Doug L. James, Jiajun Wu
摘要
Objects make unique sounds under different perturbations, environment conditions, and poses relative to the listener. While prior works have modeled impact sounds and sound propagation in simulation, we lack a standard dataset of impact sound fields of real objects for audiovisual learning and calibration of the sim-to-real gap. We present REALIMPACT, a large-scale dataset of real object impact sounds recorded under controlled conditions. RE-ALIMPACT contains 150,000 recordings of impact sounds of 50 everyday objects with detailed annotations, including their impact locations, microphone locations, contact force profiles, material labels, and RGBD images. * We make preliminary attempts to use our dataset as a reference to current simulation methods for estimating object impact sounds that match the real world. Moreover, we demonstrate the usefulness of our dataset as a testbed for acoustic and audio-visual learning via the evaluation of two benchmark tasks, including listener location classification and visual acoustic matching.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference TasksXutong Jin, Chenxi Xu, Ruohan Gao, Jiajun Wu 等SIGGRAPH 2024 · 被引用 3 次
- X-Capture: An Open-Source Portable Device for Multi-Sensory LearningSamuel Clarke, Suzannah Wistreich, Yanjie Ze, Jiajun WuICCV 2025 · 被引用 1 次
- MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D WorldYining Hong, Zishuo Zheng, Peihao Chen, Yian Wang 等CVPR 2024
它引用的顶会 Paper12
- Dual Attention Matching for Audio-Visual Event LocalizationYu Wu, Linchao Zhu, Yan Yan, Yi YangICCV 2019 · 被引用 233 次
- Discriminative Sounding Objects Localization via Self-supervised Audiovisual MatchingDi Hu, Rui Qian, Minyue Jiang, Xiao Tan 等NeurIPS 2020 · 被引用 156 次
- See, Hear, Explore: Curiosity via Audio-Visual AssociationVictoria Dean, Shubham Tulsiani, Abhinav GuptaNeurIPS 2020 · 被引用 66 次
- ObjectFolder 2.0: A Multisensory Object Dataset for Sim2Real TransferRuohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke 等CVPR 2022 · 被引用 58 次
- Audio-Visual Floorplan ReconstructionSenthil Purushwalkam, Sebastià Vicenc Amengual Garí, Vamsi Krishna Ithapu, Carl Schissler 等ICCV 2021 · 被引用 45 次
相关 Paper
- Real Acoustic Fields: An Audio-Visual Room Acoustics Dataset and BenchmarkZiyang Chen, Israel D. Gebru, Christian Richardt, Anurag Kumar 等CVPR 2024
- Finding Fallen Objects Via Asynchronous Audio-Visual IntegrationChuang Gan, Yi Gu, Siyuan Zhou, Jeremy Schwartz 等CVPR 2022 · 被引用 13 次
- Physics-Driven Diffusion Models for Impact Sound Synthesis from VideosKun Su, Kaizhi Qian, Eli Shlizerman, Antonio Torralba 等CVPR 2023
- Deep-Modal: Real-Time Impact Sound Synthesis for Arbitrary ShapesXutong Jin, Sheng Li, Tianshu Qu, Dinesh Manocha 等ACM MM 2020 · 被引用 20 次
- GWA: A Large High-Quality Acoustic Dataset for Audio ProcessingZhenyu Tang, Rohith Aralikatti, Anton Jeran Ratnarajah, Dinesh ManochaSIGGRAPH 2022 · 被引用 23 次
