Hilbert Curve-Encoded Rotation-Equivariant Oriented Object Detector with Locality-Preserving Spatial Mapping
Qi Ming, Liuqian Wang, Juan Fang, Xudong Zhao, Yucheng Xu, Ziyi Teng, Yue Zhou, Xiaoxi Hu, Xiaohan Zhang, Yufei Guo
Abstract
Arbitrary-Oriented Object Detection (AOOD) has found broad applications in embodied intelligence, autonomous driving, and satellite remote sensing. However, current AOOD frameworks face challenges in ineffective feature extraction and orientation regression inaccuracy. Inspired by Hilbert curve's intrinsic locality-preserving property, we propose a flexible Hilbert curve-Encoded Rotation-Equivariant Oriented Object Detector (HERO-Det). Our innovations include: (i) a novel Hilbert curve traversal convolution paradigm with a dimensionality reduction scheme, which employs locality-preserving spatial filling curves for feature transformation, (ii) a Hilbert pyramid transformer enabling hierarchical construction of multi-scale feature sequences through space-folding operations, as well as (iii) an orientation-adaptive prediction head that decouples rotation-equivariant regression features from invariant classification cues to resolve orientation regression dilemmas in two-stage detectors. Extensive experiments show HERO-Det achieves state-of-the-art performance on AOOD benchmarks, with mAP of 79.56%, 90.64%, 90.10%, and 80.47% on DOTA, HRSC2016, SSDD, and HRSID, respectively. Performance gains in cross-task validation further demonstrate the versatility of our method to diverse vision tasks, such as medical image segmentation and 3D object detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d043045-2bab-4344-843d-f8ee969aedb7Builds on17
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu et al.NeurIPS 2024 · 3,199 citations
- Oriented R-CNN for Object DetectionXingxing Xie, Gong Cheng, Jiabao Wang, Xiwen Yao et al.ICCV 2021 · 1,070 citations
- SCRDet: Towards More Robust Detection for Small, Cluttered and Rotated ObjectsXue Yang, Jirui Yang, Junchi Yan, Yue Zhang et al.ICCV 2019 · 865 citations
Related papers
- FRED: Towards a Full Rotation-Equivariance in Aerial Image Object DetectionChanho Lee, Jinsu Son, Hyounguk Shon, Yunho Jeon et al.AAAI 2024 · 30 citations
- Fourier Angle Alignment for Oriented Object Detection in Remote SensingChangyu Gu, Linwei Chen, Lin Gu, Ying FuCVPR 2026 · 9 citations
- OSKDet: Orientation-sensitive Keypoint Localization for Rotated Object DetectionDongchen Lu, Dongmei Li, Yali Li, Shengjin WangCVPR 2022 · 25 citations
- Spatial Transform Decoupling for Oriented Object DetectionHongtian Yu, Yunjie Tian, Qixiang Ye, Yunfan LiuAAAI 2024 · 56 citations
- Adaptive Rotated Convolution for Rotated Object DetectionYifan Pu, Yiru Wang, Zhuofan Xia, Yizeng Han et al.ICCV 2023 · 154 citations
