SE(3) Equivariant Convolution and Transformer in Ray Space
Yinshuang Xu, Jiahui Lei, Kostas Daniilidis
摘要
3D reconstruction and novel view rendering can greatly benefit from geometric priors when the input views are not sufficient in terms of coverage and inter-view baselines. Deep learning of geometric priors from 2D images requires each image to be represented in a 2D canonical frame and the prior to be learned in a given or learned 3D canonical frame. In this paper, given only the relative poses of the cameras, we show how to learn priors from multiple views equivariant to coordinate frame transformations by proposing an SE(3)-equivariant convolution and transformer in the space of rays in 3D. We model the ray space as a homogeneous space of SE(3) and introduce the SE(3)-equivariant convolution in ray space. Depending on the output domain of the convolution, we present convolution-based SE(3)-equivariant maps from ray space to ray space and to R 3 . Our mathematical framework allows us to go beyond convolution to SE(3)-equivariant attention in the ray space. We showcase how to tailor and adapt the equivariant convolution and transformer in the tasks of equivariant 3D reconstruction and equivariant neural rendering from multiple views. We demonstrate SE(3)-equivariance by obtaining robust results in roto-translated datasets without performing transformation augmentation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Equivariant Ray Embeddings for Implicit Multi-View Depth EstimationYinshuang Xu, Dian Chen, Katherine Liu, Sergey Zakharov 等NeurIPS 2024 · 被引用 11 次
- EqNIO: Subequivariant Neural Inertial OdometryRoyina Karegoudra Jayanth, Yinshuang Xu, Ziyun Wang, Evangelos Chatzipantazis 等ICLR 2025
它引用的顶会 Paper34
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 被引用 1,432 次
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima 等ICCV 2019 · 被引用 1,411 次
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 被引用 1,025 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- Light Field Networks: Neural Scene Representations with Single-Evaluation RenderingVincent Sitzmann, Semon Rezchikov, Bill Freeman, Josh Tenenbaum 等NeurIPS 2021 · 被引用 426 次
相关 Paper
- Pose-Transformed Equivariant Network for 3D Point Trajectory PredictionRuixuan Yu, Jian SunCVPR 2024 · 被引用 2 次
- SE(3)-bi-equivariant Transformers for Point Cloud AssemblyZiming Wang, Rebecka JörnstenNeurIPS 2024 · 被引用 5 次
- Gaussian Process Priors for View-Aware InferenceYuxin Hou, Ari Heljakka, Arno SolinAAAI 2021 · 被引用 1 次
- Equivariant Single View Pose Prediction Via Induced and Restriction RepresentationsOwen Howell, David Klee, Ondrej Biza, Linfeng Zhao 等NeurIPS 2023 · 被引用 4 次
- Learning Coordinate-based Convolutional Kernels for Continuous SE(3) Equivariant and Efficient Point Cloud AnalysisJaein Kim, Hee Bin Yoo, Dong-Sig Han, Byoung-Tak ZhangCVPR 2026
