GenFlow: Generalizable Recurrent Flow for 6D Pose Refinement of Novel Objects
Sungphill Moon, Hyeontae Son, Dongcheol Hur, Sangwook Kim
摘要
Despite the progress of learning-based methods for 6D object pose estimation, the tradeoff between accuracy and scalability for novel objects still exists. Specifically, previous methods for novel objects do not make good use of the target object's 3D shape information since they focus on generalization by processing the shape indirectly, making them less effective. We present GenFlow, an approach that enables both accuracy and generalization to novel objects with the guidance of the target object's shape. Our method predicts optical flow between the rendered image and the observed image and refines the 6D pose iteratively. It boosts the performance by a constraint of the 3D shape and the generalizable geometric knowledge learned from an end-to-end differentiable system. We further improve our model by designing a cascade network architecture to exploit the multi-scale correlations and coarse-to-fine refinement. GenFlow ranked first on the unseen object pose estimation benchmarks in both the RGB and RGB-D cases. It also achieves performance competitive with existing state-of-the-art methods for the seen object pose estimation without any fine-tuning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- GaussianNexus: Room-Scale Real-Time AR/VR Telepresence with Gaussian SplattingXincheng Huang, Dieter Frehlich, Ziyi Xia, Peyman Gholami 等UIST 2025 · 被引用 5 次
- Event6D: Event-based Novel Object 6D Pose TrackingJae-Young Kang, Hoonhee Cho, Taeyeop Lee, Minjun Kang 等CVPR 2026 · 被引用 4 次
- EgoXtreme: A Dataset for Robust Object Pose Estimation in Egocentric Views under Extreme ConditionsTaegyoon Yoon, Yegyu Han, Seojin Ji, Jaewoo Park 等CVPR 2026 · 被引用 3 次
- COG: Confidence-aware Optimal Geometric Correspondence for Unsupervised Single-reference Novel Object Pose EstimationYuchen Che, JINGTU WU, Hao ZHENG, Asako KanezakiCVPR 2026 · 被引用 1 次
- 3D-Object Perception Transformer (3PT)Agastya Kalra, Tim Salzmann, Guy Stoppi, Dmitrii Marin 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper15
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera 等ICCV 2019 · 被引用 504 次
- OnePose++: Keypoint-Free One-Shot Object Pose Estimation without CAD ModelsXingyi He, Jiaming Sun, Yuang Wang, Di Huang 等NeurIPS 2022 · 被引用 190 次
- EPro-PnP: Generalized End-to-End Probabilistic Perspective-n-Points for Monocular Object Pose EstimationHansheng Chen, Pichao Wang, Fan Wang, Wei Tian 等CVPR 2022 · 被引用 175 次
相关 Paper
- Shape-Constraint Recurrent Flow for 6D Object Pose EstimationYang Hai, Rui Song, Jiaojiao Li, Yinlin HuCVPR 2023
- SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene FlowQingyuan Wang, Rui Song, Jiaojiao Li, Kerui Cheng 等CVPR 2025
- Co-op: Correspondence-based Novel Object Pose EstimationSungphill Moon, Hyeontae Son, Dongcheol Hur, Sangwook KimCVPR 2025
- RefPose: Leveraging Reference Geometric Correspondences for Accurate 6D Pose Estimation of Unseen ObjectsJaeguk Kim, Jaewoo Park, Keuntek Lee, Nam Ik ChoCVPR 2025
- iG-6DoF: Model-free 6DoF Pose Estimation for Unseen Object via Iterative 3D Gaussian SplattingTuo Cao, Fei Luo, Jiongming Qin, Yu Jiang 等CVPR 2025
