SimROD: A Simple Baseline for Raw Object Detection with Global and Local Enhancements
Haiyang Xie, Xi Shen, Shihua Huang, Qirui Wang, Zheng Wang
Abstract
Most visual models are designed for sRGB images, yet RAW data offers significant advantages for object detection by preserving sensor information before ISP processing. This enables improved detection accuracy and more efficient hardware designs by bypassing the ISP. However, RAW object detection is challenging due to limited training data, unbalanced pixel distributions, and sensor noise. To address this, we propose SimROD, a lightweight and effective approach for RAW object detection. We introduce a Global Gamma Enhancement (GGE) module, which applies a learnable global gamma transformation with only four parameters, improving feature representation while keeping the model efficient. Additionally, we leverage the green channel's richer signal to enhance local details, aligning with the human eye’s sensitivity and Bayer filter design. Extensive experiments on multiple RAW object detection datasets and detectors demonstrate that SimROD outperforms state-of-the-art methods like RAW-Adapter and DIAP while maintaining efficiency. Our work highlights the potential of RAW data for real-world object detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 908bc68b-9523-4187-98bd-ab2d1ca7f4b7Cited by top-tier papers1
Ask how each one uses itBuilds on11
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- RawHDR: High Dynamic Range Image Reconstruction from a Single Raw ImageYunhao Zou, Chenggang Yan, Ying FuICCV 2023 · 36 citations
- HumanNeRF-SE: A Simple yet Effective Approach to Animate HumanNeRF with Diverse PosesCaoyuan Ma, Yu-Lun Liu, Zhixiang Wang, Wu Liu et al.CVPR 2024 · 8 citations
- You Think, You ACT: the New Task of Arbitrary Text to Motion GenerationRunqi Wang, Caoyuan Ma, Guopeng Li, Hanrui Xu et al.ICCV 2025 · 3 citations
- Guiding a Harsh-Environments Robust Detector via RAW Data Characteristic MiningHongyang Chen, Hung-Shuo Tai, Kaisheng MaAAAI 2024 · 2 citations
Related papers
- Beyond RGB: Adaptive Parallel Processing for RAW Object DetectionShani Gamrian, Hila Barel, Feiran Li, Masakazu Yoshimura et al.ICCV 2025 · 4 citations
- Task-Aware Image Signal Processor for Advanced Visual PerceptionKai Chen, Jin Xiao, Leheng Zhang, Kexuan Shi et al.CVPR 2026 · 3 citations
- Dark-ISP: Enhancing RAW Image Processing for Low-Light Object DetectionJiasheng Guo, Xin Gao, Yuxiang Yan, Guanghao Li et al.ICCV 2025 · 5 citations
- Toward RAW Object Detection: A New Benchmark and A New ModelRuikang Xu, Chang Chen, Jingyang Peng, Cheng Li et al.CVPR 2023
- SpiralDiff: Spiral Diffusion with LoRA for RGB-to-RAW Conversion Across CamerasHuanjing Yue, Shangbin Xie, Cong Cao, Qian Wu et al.CVPR 2026
