Cylindrical Convolutional Networks for Joint Object Detection and Viewpoint Estimation
Sunghun Joung, Seungryong Kim, Hanjae Kim, Minsu Kim, Ig-Jae Kim, Junghyun Cho, Kwanghoon Sohn
摘要
Existing techniques to encode spatial invariance within deep convolutional neural networks only model 2D transformation fields. This does not account for the fact that objects in a 2D space are a projection of 3D ones, and thus they have limited ability to severe object viewpoint changes. To overcome this limitation, we introduce a learnable module, cylindrical convolutional networks (CCNs), that exploit cylindrical representation of a convolutional kernel defined in the 3D space. CCNs extract a view-specific feature through a view-specific convolutional kernel to predict object category scores at each viewpoint. With the viewspecific feature, we simultaneously determine objective category and viewpoints using the proposed sinusoidal softargmax module. Our experiments demonstrate the effectiveness of the cylindrical convolutional networks on joint object detection and viewpoint estimation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Localization with Sampling-ArgmaxJiefeng Li, Tong Chen, Ruiqi Shi, Yujing Lou 等NeurIPS 2021 · 被引用 25 次
- Learning Canonical 3D Object Representation for Fine-Grained RecognitionSunghun Joung, Seungryong Kim, Minsu Kim, Ig-Jae Kim 等ICCV 2021 · 被引用 14 次
- ViewNet: Unsupervised Viewpoint Estimation from Conditional GenerationOctave Mariotti, Oisin Mac Aodha, Hakan BilenICCV 2021 · 被引用 8 次
- Uncertainty-Aware Joint Salient Object and Camouflaged Object DetectionAixuan Li, Jing Zhang, Yunqiu Lv, Bowen Liu 等CVPR 2021
- SpinNet: Learning a General Surface Descriptor for 3D Point Cloud RegistrationSheng Ao, Qingyong Hu, Bo Yang, Andrew Markham 等CVPR 2021
相关 Paper
- VI-Net: Boosting Category-level 6D Object Pose Estimation via Learning Decoupled Rotations on the Spherical RepresentationsJiehong Lin, Zewei Wei, Yabin Zhang, Kui JiaICCV 2023 · 被引用 57 次
- RRL: Regional Rotate Layer in Convolutional Neural NetworksZongbo Hao, Tao Zhang, Mingwang Chen, Kaixu ZhouAAAI 2022 · 被引用 6 次
- AeDet: Azimuth-Invariant Multi-View 3D Object DetectionChengjian Feng, Zequn Jie, Yujie Zhong, Xiangxiang Chu 等CVPR 2023
- Learning Shape-Independent Transformation via Spherical Representations for Category-Level Object Pose EstimationHuan Ren, Wenfei Yang, Xiang Liu, Shifeng Zhang 等ICLR 2025
- FisheyeHDK: Hyperbolic Deformable Kernel Learning for Ultra-Wide Field-of-View Image RecognitionOla Ahmad, Freddy LécuéAAAI 2022 · 被引用 21 次
