Towards RAW Object Detection in Diverse Conditions
Zhongyu Li, Xin Jin, Boyuan Sun, Chun-Le Guo, Ming-Ming Cheng
摘要
Existing object detection methods often consider sRGB input, which was compressed from RAW data using ISP originally designed for visualization. However, such compression might lose crucial information for detection, especially under complex light and weather conditions. We introduce the AODRaw dataset, which offers 7,785 highresolution real RAW images with 135,601 annotated instances spanning 62 categories, capturing a broad range of indoor and outdoor scenes under 9 distinct light and weather conditions. Based on AODRaw that supports RAW and sRGB object detection, we provide a comprehensive benchmark for evaluating current detection methods. We find that sRGB pre-training constrains the potential of RAW object detection due to the domain gap between sRGB and RAW, prompting us to directly pre-train on the RAW domain. However, it is harder for RAW pre-training to learn rich representations than sRGB pre-training. To assist RAW pre-training, we distill the knowledge from an off-the-shelf model pre-trained on the sRGB domain. As a result, we achieve substantial improvements under diverse and adverse conditions without relying on extra pre-processing modules. The code and dataset are available at https: //github.com/lzyhha/AODRaw .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- DitHub: A Modular Framework for Incremental Open-Vocabulary Object DetectionChiara Cappellino, Gianluca Mancusi, Matteo Mosconi, Angelo Porrello 等NeurIPS 2025 · 被引用 4 次
- End-to-End Low-Light Enhancement for Object Detection with Learned Metadata from RAWsXuelin Shen, Haifeng Jiao, Yitong Wang, Yulin He 等NeurIPS 2025 · 被引用 1 次
- Bridging RGB and RAW: Single-step Deterministic Flow with Homogeneous Representation AlignmentDiedong Feng, Peiyi Zeng, Zhen Liu, Zhongyang Li 等ICML 2026
- SpiralDiff: Spiral Diffusion with LoRA for RGB-to-RAW Conversion Across CamerasHuanjing Yue, Shangbin Xie, Cong Cao, Qian Wu 等CVPR 2026
它引用的顶会 Paper16
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object DetectionXiang Li, Wenhai Wang, Lijun Wu, Shuo Chen 等NeurIPS 2020 · 被引用 2,118 次
- ReconfigISP: Reconfigurable Camera Image Processing PipelineKe Yu, Zexian Li, Yue Peng, Chen Change Loy 等ICCV 2021 · 被引用 46 次
相关 Paper
- Guiding a Harsh-Environments Robust Detector via RAW Data Characteristic MiningHongyang Chen, Hung-Shuo Tai, Kaisheng MaAAAI 2024 · 被引用 2 次
- Adaptive Illumination Mapping for Shadow Detection in Raw ImagesJiayu Sun, Ke Xu, Youwei Pang, Lihe Zhang 等ICCV 2023 · 被引用 15 次
- SimROD: A Simple Baseline for Raw Object Detection with Global and Local EnhancementsHaiyang Xie, Xi Shen, Shihua Huang, Qirui Wang 等AAAI 2026 · 被引用 1 次
- ReRAW: RGB-to-RAW Image Reconstruction via Stratified Sampling for Efficient Object Detection on the EdgeRadu Berdan, Beril Besbinar, Christoph Reinders, Junji Otsuka 等CVPR 2025
- Toward RAW Object Detection: A New Benchmark and A New ModelRuikang Xu, Chang Chen, Jingyang Peng, Cheng Li 等CVPR 2023
