InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
Jinlu Zhang, Yixin Chen, Zan Wang, Jie Yang, Yizhou Wang, Siyuan Huang
Abstract
Diverse human interactions. (b) Detailed interactions synthesis of open-set objects. (c) Novel interactions given generative objects. A person holds/pulls/lies on the chair… A person grasps the backpack/table/ball/knife… A person lifts/rides/craddles the tesla car/motorcycle/horse/baby… Figure 1. 3D human object interaction synthesis by InteractAnything. Given a simple text description with goal interaction and any object mesh as input, our method enables the generation of diverse, natural, detailed, and novel interactions for open-set 3D objects in a zero-shot manner. The orange and green boxes of (b) indicate detailed contact poses from different views.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf268ae5-83b4-4016-b9ec-cd8cbfd89ab8Cited by top-tier papers4
- InterPrior: Scaling Generative Control for Physics-Based Human-Object InteractionsSirui Xu, Samuel Schulter, Morteza Ziyadi, Xialin He et al.CVPR 2026 · 14 citations
- Decoupled Generative Modeling for Human-Object Interaction SynthesisHwanhee Jung, Seunggwan Lee, Jeongyoon Yoon, SeungHyeon Kim et al.CVPR 2026 · 4 citations
- H2OFlow: Grounding Human-Object Affordances with 3D Generative Models and Dense Diffused FlowsHarry Zhang, Luca CarloneICLR 2026 · 2 citations
- HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance GuidanceLei Li, Angela DaiICML 2026
Builds on33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
Related papers
- CG-HOI: Contact-Guided 3D Human-Object Interaction GenerationChristian Diller, Angela DaiCVPR 2024
- GenZI: Zero-Shot 3D Human-Scene Interaction GenerationLei Li, Angela DaiCVPR 2024
- Text2HOI: Text-Guided 3D Motion Generation for Hand-Object InteractionJunuk Cha, Jihyeon Kim, Jae Shin Yoon, Seungryul BaekCVPR 2024
- ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction GenerationLing-An Zeng, Guohong Huang, Yi-Lin Wei, Shengbo Gu et al.CVPR 2025
- InterDiff: Generating 3D Human-Object Interactions with Physics-Informed DiffusionSirui Xu, Zhengyuan Li, Yu-Xiong Wang, Liang-Yan GuiICCV 2023 · 201 citations
