An Analysis of SVD for Deep Rotation Estimation
Jake Levinson, Carlos Esteves, Kefan Chen, Noah Snavely, Angjoo Kanazawa, Afshin Rostamizadeh, Ameesh Makadia
摘要
Symmetric orthogonalization via SVD, and closely related procedures, are well-known techniques for projecting matrices onto or . These tools have long been used for applications in computer vision, for example optimal 3D alignment problems solved by orthogonal Procrustes, rotation averaging, or Essential matrix decomposition. Despite its utility in different settings, SVD orthogonalization as a procedure for producing rotation matrices is typically overlooked in deep learning models, where the preferences tend toward classic representations like unit quaternions, Euler angles, and axis-angle, or more recently-introduced methods. Despite the importance of 3D rotations in computer vision and robotics, a single universally effective representation is still missing. Here, we explore the viability of SVD orthogonalization for 3D rotations in neural networks. We present a theoretical analysis that shows SVD is the natural choice for projecting onto the rotation group. Our extensive quantitative analysis shows simply replacing existing representations with the SVD orthogonalization procedure obtains state of the art performance in many deep learning applications covering both supervised and unsupervised training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Implicit-PDF: Non-Parametric Representation of Probability Distributions on the Rotation ManifoldKieran A. Murphy, Carlos Esteves, Varun Jampani, Srikumar Ramalingam 等ICML 2021 · 被引用 93 次
- SAR-Net: Shape Alignment and Recovery Network for Category-level 6D Object Pose and Size EstimationHaitao Lin, Zichang Liu, Chilam Cheang, Yanwei Fu 等CVPR 2022 · 被引用 86 次
- Learning with 3D rotations, a hitchhiker's guide to SO(3)Andreas René Geist, Jonas Frey, Mikel Zhobro, Anna Levina 等ICML 2024 · 被引用 47 次
- Disentangled3D: Learning a 3D Generative Model with Disentangled Geometry and Appearance from Monocular ImagesAyush Tewari, Mallikarjun B. R., Xingang Pan, Ohad Fried 等CVPR 2022 · 被引用 35 次
- Projective Manifold Gradient Layer for Deep Rotation RegressionJiayi Chen, Yingda Yin, Tolga Birdal, Baoquan Chen 等CVPR 2022 · 被引用 16 次
它引用的顶会 Paper1
相关 Paper
- Eliminating topological errors in neural network rotation estimation using self-selecting ensemblesSitao XiangSIGGRAPH 2021 · 被引用 7 次
- Vector Neurons: A General Framework for SO(3)-Equivariant NetworksCongyue Deng, Or Litany, Yueqi Duan, Adrien Poulenard 等ICCV 2021 · 被引用 411 次
- Quaternion Product Units for Deep Learning on 3D Rotation GroupsXuan Zhang, Shaofei Qin, Yi Xu, Hongteng XuCVPR 2020
- Equivariant Single View Pose Prediction Via Induced and Restriction RepresentationsOwen Howell, David Klee, Ondrej Biza, Linfeng Zhao 等NeurIPS 2023 · 被引用 4 次
- Learning to Orient Surfaces by Self-supervised Spherical CNNsRiccardo Spezialetti, Federico Stella, Marlon Marcon, Luciano Silva 等NeurIPS 2020 · 被引用 48 次
