Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
Jingyu Gong, Kunkun Tong, Zhuoran Chen, Chuanhan Yuan, Mingang Chen, Zhizhong Zhang, Xin Tan, Yuan Xie
摘要
Human motion synthesis in 3D scenes relies heavily on scene comprehension, while current methods focus mainly on scene structure but ignore the semantic understanding. In this paper, we propose a human motion synthesis framework that take an unified Scene Semantic Occupancy (SSO) for scene representation, termed SSOMotion. We design a bi-directional triplane decomposition to derive a compact version of the SSO, and scene semantics are mapped to an unified feature space via CLIP encoding and shared linear dimensionality reduction. Such strategy can derive the fine-grained scene semantic structures while significantly reduce redundant computations. We further take these scene hints and movement direction derived from instructions for motion control via framewise scene query. Extensive experiments and ablation studies conducted on cluttered scenes using ShapeNet furniture, as well as scanned scenes from PROX and Replica datasets, demonstrate its cutting-edge performance while validating its effectiveness and generalization ability. Code will be publicly available at https://github.com/jingyugong/SSOMotion .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative LearningZhenwu Shi, Jingyu Gong, Peiwei Wang, Xingzan Wang 等CVPR 2026 · 被引用 4 次
- Open the Motion Door: Atomic Motion Decomposition and Recomposition for Open-Vocabulary Motion GenerationKe Fan, Jiangning Zhang, Ran Yi, Jingyu Gong 等CVPR 2026
它引用的顶会 Paper24
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 被引用 672 次
- Generating Diverse and Natural 3D Human Motions from TextChuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang 等CVPR 2022 · 被引用 462 次
- Resolving 3D Human Pose Ambiguities With 3D Scene ConstraintsMohamed Hassan, Vasileios Choutas, Dimitrios Tzionas, Michael J. BlackICCV 2019 · 被引用 384 次
相关 Paper
- Diffusion Implicit Policy for Unpaired Scene-aware Motion SynthesisJingyu Gong, Chong Zhang, Fengqi Liu, Ke Fan 等AAAI 2026 · 被引用 5 次
- Synthesizing Long-Term 3D Human Motion and Interaction in 3D ScenesJiashun Wang, Huazhe Xu, Jingwei Xu, Sifei Liu 等CVPR 2021
- COAP: Compositional Articulated Occupancy of PeopleMarko Mihajlovic, Shunsuke Saito, Aayush Bansal, Michael Zollhöfer 等CVPR 2022 · 被引用 45 次
- MIME: Human-Aware 3D Scene GenerationHongwei Yi, Chun-Hao P. Huang, Shashank Tripathi, Lea Hering 等CVPR 2023
- S2GO: Streaming Sparse Gaussian OccupancyJinhyung Park, Chensheng Peng, Yihan Hu, Wenzhao Zheng 等ICLR 2026 · 被引用 2 次
