Scalable RF Simulation in Generative 4D Worlds
Zhiwei Zheng, Dongyin Hu, Mingmin Zhao
Abstract
Radio Frequency (RF) sensing has emerged as a powerful, privacy-preserving alternative to visionbased methods for various perception tasks. However, building high-quality RF datasets in dynamic and diverse environments remains a major challenge. To address this, we introduce WAVEV-ERSE, a prompt-based, scalable framework that simulates realistic RF signals from generated indoor scenes with human motions guided by spatial paths, enabling diverse and feasible behaviors without manual trajectory design. WAVEV-ERSE features a language-guided 4D world generator and a physics-based signal simulator that enables realistic simulation of RF signals in diverse environments. It employs a phase-coherent ray tracer that preserves both spatial and temporal phase consistency. The simulated signals show high fidelity on phase-sensitive benchmarks, and closely align with both real-world collected measurements and simulations from a proprietary electromagnetic solver. When used for data augmentation, WAVEVERSE consistently improves performance in downstream tasks like RF imaging and human activity recognition, with gains that grow with the amount of simulated data and surpass existing methods. Code and additional materials are available on the webpage.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- SmartDJ: Declarative Audio Editing with Audio Language ModelZitong Lan, Yiduo Hao, Mingmin ZhaoICLR 2026 · 11 citations
- Next-Scale Autoregressive Models for Text-to-Motion GenerationZhiwei Zheng, Shibo Jin, Lingjie Liu, Mingmin ZhaoCVPR 2026 · 6 citations
Builds on19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- MotionGPT: Human Motion as a Foreign LanguageBiao Jiang, Xin Chen, Wen Liu, Jingyi Yu et al.NeurIPS 2023 · 698 citations
- Generating Diverse and Natural 3D Human Motions from TextChuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang et al.CVPR 2022 · 462 citations
- Human Motion Diffusion as a Generative PriorYoni Shafir, Guy Tevet, Roy Kapon, Amit Haim BermanoICLR 2024 · 371 citations
Related papers
- RF-protect: privacy against device-free human trackingJayanth Shenoy, Zikun Liu, Bill Tao, Zachary Kabelac et al.SIGCOMM 2022 · 42 citations
- Towards Generalized mmWave-based Human Pose Estimation through Signal AugmentationHongfei Xue, Qiming Cao, Chenglin Miao, Yan Ju et al.MobiCom 2023 · 69 citations
- M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh ReconstructionJunqiao Fan, Yunjiao Zhou, Yizhuo Yang, Xinyuan Cui et al.CVPR 2026 · 11 citations
- One Snapshot is All You Need: A Generalized Method for mmWave Signal GenerationTeng Huang, Han Ding, Wenxin Sun, Cui Zhao et al.INFOCOM 2025 · 6 citations
- Wave-Former: Through-Occlusion 3D Reconstruction via Wireless Shape CompletionLaura Dodds, Maisy Lam, Waleed Akbar, Yibo Cheng et al.CVPR 2026
