OW-DAR: Dual-Granularity Adaptive Reconstruction-Error Modeling for Open-World Object Detection
Linhua Ye, Xing Xi, Ronghua Luo
Abstract
Open-world object detection (OWOD) aims to detect known and unknown objects in dynamic environments. However, only known classes are labeled during training, making it challenging for detectors to recognize unknown objects during inference. Existing methods typically rely on supervision from known categories, leading models to overconfidently misclassify visually similar unknowns as known, and dissimilar ones as background. This known-class prior bias limits the model's ability to detect unknown objects. In this paper, we propose a novel method, OW-DAR, which enhances foreground-background separability through collaborative fine-grained and coarse-grained modeling. At the finegrained level, we propose Fine-grained Masked Reconstruction (FMR), which randomly masks regions of the feature map to guide the reconstruction toward semantic structures, rather than memorizing low-level patterns. At the coarsegrained level, we propose Adaptive Region-based Error Aggregation (AREA), which operates on object proposals to aggregate reconstruction errors. This enables the model to attend to semantically ambiguous foreground-background boundaries while suppressing the influence of local outliers during optimization. Finally, we leverage robust reconstruction errors to perform unsupervised foreground-background modeling, enabling probabilistic estimation for potential unknown objects. We validate the effectiveness of OW-DAR on standard OWOD benchmark. Experimental results demonstrate that OW-DAR consistently outperforms existing stateof-the-art methods, achieving a +18.8 improvement in unknown object recall (U-Recall).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- DETRs Beat YOLOs on Real-time Object DetectionYian Zhao, Wenyu Lv, Shangliang Xu, Jinman Wei et al.CVPR 2024 · 3,046 citations
- OW-DETR: Open-world Detection TransformerAkshita Gupta, Sanath Narayan, K. J. Joseph, Salman Khan et al.CVPR 2022 · 209 citations
- FreeSOLO: Learning to Segment Objects without AnnotationsXinlong Wang, Zhiding Yu, Shalini De Mello, Jan Kautz et al.CVPR 2022 · 100 citations
- Rethinking Reconstruction Autoencoder-Based Out-of-Distribution DetectionYibo ZhouCVPR 2022 · 69 citations
Related papers
- UMB: Understanding Model Behavior for Open-World Object DetectionXing Xi, Yangyang Huang, Zhijie Zhong, Ronghua LuoNeurIPS 2024 · 10 citations
- PROB: Probabilistic Objectness for Open World Object DetectionOrr Zohar, Kuan-Chieh Wang, Serena YeungCVPR 2023
- Open-World Objectness Modeling Unifies Novel Object DetectionShan Zhang, Yao Ni, Jinhao Du, Yuan Xue et al.CVPR 2025
- OW-Adapter: Human-Assisted Open-World Object Detection with a Few ExamplesSuphanut Jamonnak, Jiajing Guo, Wenbin He, Liang Gou et al.IEEE VIS 2023 · 6 citations
- Knowing the Unknown: Interpretable Open-World Object Detection via Concept Decomposition ModelXueqiang Lv, Shizhou Zhang, Yinghui Xing, di xu et al.ICML 2026 · 2 citations
