IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
Can Zhang, Gim Hee Lee
摘要
This work presents IAAO, a novel framework that builds an explicit 3D model for intelligent agents to gain understanding of articulated objects in their environment through interaction. Unlike prior methods that rely on task-specific networks and assumptions about movable parts, our IAAO leverages large foundation models to estimate interactive affordances and part articulations in three stages. We first build hierarchical features and label fields for each object state using 3D Gaussian Splatting (3DGS) by distilling mask features and view-consistent labels from multi-view images. We then perform object-and part-level queries on the 3D Gaussian primitives to identify static and articulated elements, estimating global transformations and local articulation parameters along with affordances. Finally, scenes from different states are merged and refined based on the estimated transformations, enabling robust affordancebased interaction and manipulation of objects. Experimental results demonstrate the effectiveness of our method. Our source code is available at: https://lulusindazc. github.io/IAAOproject/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Clay-to-Stone: Phase-wise 3D Gaussian Splatting for Monocular Articulated Hand-Object Manipulation ModelingXingyu Liu, Pengfei Ren, Qi Qi, Haifeng Sun 等CVPR 2026 · 被引用 1 次
- ArtPro: Self-Supervised Articulated Object Reconstruction with Adaptive Integration of Mobility ProposalsXuelu Li, Zhaonan Wang, Xiaogang Wang, Lei Wu 等CVPR 2026 · 被引用 1 次
- SimArt: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLMChuanrui Zhang, Minghan Qin, Yuang Wang, Baifeng Xie 等SIGGRAPH 2026
- LAM: Language Articulated Object ModelersYipeng Gao, Yunhao Ge, Peilin Cai, Daniel Seita 等CVPR 2026
- SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D ObservationCan Zhang, Gim Hee LeeCVPR 2026
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Vision Transformers Need RegistersTimothée Darcet, Maxime Oquab, Julien Mairal, Piotr BojanowskiICLR 2024 · 被引用 769 次
- LERF: Language Embedded Radiance FieldsJustin Kerr, Chung Min Kim, Ken Goldberg, Angjoo Kanazawa 等ICCV 2023 · 被引用 620 次
相关 Paper
- 3DAffordSplat: Efficient Affordance Reasoning with 3D GaussiansZeming Wei, Junyi Lin, Yang Liu, Weixing Chen 等ACM MM 2025 · 被引用 4 次
- SPLART: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian SplattingShengjie Lin, Jiading Fang, Muhammad Zubair Irshad, Vitor Campagnolo Guizilini 等ICCV 2025 · 被引用 2 次
- Building Interactable Replicas of Complex Articulated Objects via Gaussian SplattingYu Liu, Baoxiong Jia, Ruijie Lu, Junfeng Ni 等ICLR 2025
- ArticulatedGS: Self-supervised Digital Twin Modeling of Articulated Objects using 3D Gaussian SplattingJunfu Guo, Yu Xin, Gaoyi Liu, Kai Xu 等CVPR 2025
- REALM: An MLLM-Agent Framework for Open World 3D Reasoning Segmentation and Editing on Gaussian SplattingChangyue Shi, Minghao Chen, Yiping Mao, Chuxiao Yang 等CVPR 2026 · 被引用 8 次
