CARTO: Category and Joint Agnostic Reconstruction of ARTiculated Objects
Nick Heppert, Muhammad Zubair Irshad, Sergey Zakharov, Katherine Liu, Rares Andrei Ambrus, Jeannette Bohg, Abhinav Valada, Thomas Kollar
摘要
We present CARTO, a novel approach for reconstructing multiple articulated objects from a single stereo RGB observation. We use implicit object-centric representations and learn a single geometry and articulation decoder for multiple object categories. Despite training on multiple categories, our decoder achieves a comparable reconstruction accuracy to methods that train bespoke decoders separately for each category. Combined with our stereo image encoder we infer the 3D shape, 6D pose, size, joint type, and the joint state of multiple unknown objects in a single forward pass. Our method achieves a 20.4% absolute improvement in mAP 3D IOU50 for novel instances when compared to a two-stage pipeline. Inference time is fast and can run on a NVIDIA TITAN XP GPU at 1 HZ for eight or less objects present. While only trained on simulated data, CARTO transfers to real-world object instances. Code and evaluation data is available at: carto. cs. uni - freiburg. de
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- PARIS: Part-level Reconstruction and Motion Analysis for Articulated ObjectsJiayi Liu, Ali Mahdavi-Amiri, Manolis SavvaICCV 2023 · 被引用 103 次
- Where2Explore: Few-shot Affordance Learning for Unseen Novel Categories of Articulated ObjectsChuanruo Ning, Ruihai Wu, Haoran Lu, Kaichun Mo 等NeurIPS 2023 · 被引用 64 次
- NeO 360: Neural Fields for Sparse View Synthesis of Outdoor ScenesMuhammad Zubair Irshad, Sergey Zakharov, Katherine Liu, Vitor Guizilini 等ICCV 2023 · 被引用 63 次
- Articulate your NeRF: Unsupervised articulated object modeling via conditional view synthesisJianning Deng, Kartic Subr, Hakan BilenNeurIPS 2024 · 被引用 26 次
- URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language ModelZhe Li, Xiang Bai, Jieyu Zhang, Zhuangzhe Wu 等NeurIPS 2025 · 被引用 24 次
它引用的顶会 Paper16
- A-NeRF: Articulated Neural Radiance Fields for Learning Human Shape, Appearance, and PoseShih-Yang Su, Frank Yu, Michael Zollhöfer, Helge RhodinNeurIPS 2021 · 被引用 316 次
- A-SDF: Learning Disentangled Signed Distance Functions for Articulated Shape RepresentationJiteng Mu, Weichao Qiu, Adam Kortylewski, Alan L. Yuille 等ICCV 2021 · 被引用 138 次
- CAPTRA: CAtegory-level Pose Tracking for Rigid and Articulated Objects from Point CloudsYijia Weng, He Wang, Qiang Zhou, Yuzhe Qin 等ICCV 2021 · 被引用 119 次
- Ditto: Building Digital Twins of Articulated Objects from InteractionZhenyu Jiang, Cheng-Chun Hsu, Yuke ZhuCVPR 2022 · 被引用 77 次
- AKB-48: A Real-World Articulated Object Knowledge BaseLiu Liu, Wenqiang Xu, Haoyuan Fu, Sucheng Qian 等CVPR 2022 · 被引用 64 次
相关 Paper
- FroDO: From Detections to 3D ObjectsMartin Rünz, Kejie Li, Meng Tang, Lingni Ma 等CVPR 2020
- ART: Articulated Reconstruction TransformerZizhang Li, Cheng Zhang, Zhengqin Li, Henry Howard-Jenkins 等CVPR 2026 · 被引用 12 次
- Detection Based Part-level Articulated Object Reconstruction from Single RGBD ImageYuki Kawana, Tatsuya HaradaNeurIPS 2023 · 被引用 20 次
- Multi-Path Learning for Object Pose Estimation Across DomainsMartin Sundermeyer, Maximilian Durner, En Yen Puang, Zoltan-Csaba Marton 等CVPR 2020
- R^2-Art: Category-Level Articulation Pose Estimation from Single RGB Image via Cascade Render StrategyLi Zhang, Haonan Jiang, Yukang Huo, Yan Zhong 等AAAI 2025 · 被引用 6 次
