FrameNet: Learning Local Canonical Frames of 3D Surfaces From a Single RGB Image
Jingwei Huang, Yichao Zhou, Thomas A. Funkhouser, Leonidas J. Guibas
摘要
In this work, we introduce the novel problem of identifying dense canonical 3D coordinate frames from a single RGB image. We observe that each pixel in an image corresponds to a surface in the underlying 3D geometry, where a canonical frame can be identified as represented by three orthogonal axes, one along its normal direction and two in its tangent plane. We propose an algorithm to predict these axes from RGB. Our first insight is that canonical frames computed automatically with recently introduced direction field synthesis methods can provide training data for the task. Our second insight is that networks designed for surface normal prediction provide better results when trained jointly to predict canonical frames, and even better when trained to also predict 2D projections of canonical frames. We conjecture this is because projections of canonical tangent directions often align with local gradients in images, and because those directions are tightly linked to 3D canonical frames through projective geometry and orthogonality constraints. In our experiments, we find that our method predicts 3D canonical frames that can be used in applications ranging from surface normal estimation, feature matching, and augmented reality.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Transformer-Based Attention Networks for Continuous Pixel-Wise PredictionGuanglei Yang, Hao Tang, Mingli Ding, Nicu Sebe 等ICCV 2021 · 被引用 246 次
- Estimating and Exploiting the Aleatoric Uncertainty in Surface Normal EstimationGwangbin Bae, Ignas Budvytis, Roberto CipollaICCV 2021 · 被引用 154 次
- UNIFIED-IO: A Unified Model for Vision, Language, and Multi-modal TasksJiasen Lu, Christopher Clark, Rowan Zellers, Roozbeh Mottaghi 等ICLR 2023 · 被引用 110 次
- Shape from Polarization for Complex Scenes in the WildChenyang Lei, Chenyang Qi, Jiaxin Xie, Na Fan 等CVPR 2022 · 被引用 60 次
- Unified-IO 2: Scaling Autoregressive Multimodal Models with Vision, Language, Audio, and ActionJiasen Lu, Christopher Clark, Sangho Lee, Zichen Zhang 等CVPR 2024 · 被引用 53 次
相关 Paper
- Canonical Fields: Self-Supervised Learning of Pose-Canonicalized Neural FieldsRohith Agaram, Shaurya Dewan, Rahul Sajnani, Adrien Poulenard 等CVPR 2023
- Symmetry-Robust 3D Orientation EstimationChristopher Scarvelis, David Ben-Haim, Paul ZhangICML 2025
- PFCNN: Convolutional Neural Networks on 3D Surfaces Using Parallel FramesYuqi Yang, Shilin Liu, Hao Pan, Yang Liu 等CVPR 2020
- Beyond Canonicalization: How Tensorial Messages Improve Equivariant Message PassingPeter Lippmann, Gerrit Gerhartz, Roman Remme, Fred A. HamprechtICLR 2025
- DualPM: Dual Posed-Canonical Point Maps for 3D Shape and Pose ReconstructionBen Kaye, Tomas Jakab, Shangzhe Wu, Christian Ruprecht 等CVPR 2025
