Primitive-Based 3D Human-Object Interaction Modelling and Programming
Siqi Liu, Yong-Lu Li, Zhou Fang, Xinpeng Liu, Yang You, Cewu Lu
Abstract
Embedding Human and Articulated Object Interaction (HAOI) in 3D is an important direction for a deeper human activity understanding. Different from previous works that use parametric and CAD models to represent humans and objects, in this work, we propose a novel 3D geometric primitive-based language to encode both humans and objects. Given our new paradigm, humans and objects are all compositions of primitives instead of heterogeneous entities. Thus, mutual information learning may be achieved between the limited 3D data of humans and different object categories. Moreover, considering the simplicity of the expression and the richness of the information it contains, we choose the superquadric as the primitive representation. To explore an effective embedding of HAOI for the machine, we build a new benchmark on 3D HAOI consisting of primitives together with their images and propose a task requiring machines to recover 3D HAOI using primitives from images. Moreover, we propose a baseline of single-view 3D reconstruction on HAOI. We believe this primitive-based 3D HAOI representation would pave the way for 3D HAOI studies. Our code and data are available at https://mvig-rhos.com/p3haoi.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 77009fc1-a977-45df-84cb-4a901928e425Cited by top-tier papers5
- SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction ScenariosLingwei Dang, Ruizhi Shao, Hongwen Zhang, Wei Min et al.NeurIPS 2025 · 12 citations
- TriDi: Trilateral Diffusion of 3D Humans, Objects, and InteractionsIlia A. Petrov, Riccardo Marin, Julian Chibane, Gerard Pons-MollICCV 2025 · 1 citation
- Reconstructing In-the-Wild Open-Vocabulary Human-Object InteractionsBoran Wen, Dingbang Huang, Zichen Zhang, Jiahong Zhou et al.CVPR 2025
- Design2GarmentCode: Turning Design Concepts to Tangible Garments Through Program SynthesisFeng Zhou, Ruiyang Liu, Chen Liu, Gaofeng He et al.CVPR 2025
- CORE4D: A 4D Human-Object-Human Interaction Dataset for Collaborative Object REarrangementYun Liu, Chengwen Zhang, Ruofan Xing, Bingda Tang et al.CVPR 2025
Builds on18
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang et al.ICCV 2021 · 376 citations
- BEHAVE: Dataset and Method for Tracking Human Object InteractionsBharat Lal Bhatnagar, Xianghui Xie, Ilya A. Petrov, Cristian Sminchisescu et al.CVPR 2022 · 144 citations
- Fully Convolutional Mesh Autoencoder using Efficient Spatially Varying KernelsYi Zhou, Chenglei Wu, Zimo Li, Chen Cao et al.NeurIPS 2020 · 98 citations
- LASSIE: Learning Articulated Shapes from Sparse Image Ensemble via 3D Part DiscoveryChun-Han Yao, Wei-Chih Hung, Yuanzhen Li, Michael Rubinstein et al.NeurIPS 2022 · 83 citations
- AKB-48: A Real-World Articulated Object Knowledge BaseLiu Liu, Wenqiang Xu, Haoyuan Fu, Sucheng Qian et al.CVPR 2022 · 64 citations
Related papers
- Detailed 2D-3D Joint Representation for Human-Object InteractionYong-Lu Li, Xinpeng Liu, Han Lu, Shiyi Wang et al.CVPR 2020
- HOI-M3: Capture Multiple Humans and Objects Interaction within Contextual EnvironmentJuze Zhang, Jingyan Zhang, Zining Song, Zhanhe Shi et al.CVPR 2024 · 7 citations
- PICO: Reconstructing 3D People In Contact with ObjectsAlpár Cseke, Shashank Tripathi, Sai Kumar Dwivedi, Arjun S. Lakshmipathy et al.CVPR 2025
- Interacted Object Grounding in Spatio-Temporal Human-Object InteractionsXiaoyang Liu, Boran Wen, Xinpeng Liu, Zizheng Zhou et al.AAAI 2025 · 6 citations
- Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using SuperquadricsTze Ho Elden Tse, Runyang Feng, Linfang Zheng, Jiho Park et al.AAAI 2025 · 2 citations
