Dr. RAW: Towards General High-Level Vision from RAW with Efficient Task Conditioning
Wenjun Huang, Ziteng Cui, Yinqiang Zheng, Yirui He, Tatsuya Harada, Mohsen Imani
Abstract
We introduce Dr. RAW , a unified and tuning-efficient framework for high-level computer vision tasks directly operating on camera RAW data. Unlike previous approaches that optimize image signal processing (ISP) pipelines and fully fine-tune networks for each task, Dr. RAW achieves state-of-the-art performance with minimal parameter updates and frozen backbone weights. At the input stage, we apply lightweight pre-processing steps, including sensor and illumination mapping, along with re-mosaicing, to mitigate data inconsistencies stemming from sensor variations and lighting conditions. At the network level, we introduce task-specific adaptation through two modules: Sensor Prior Prompts (SPP) and task-specific Low-Rank Adaptation (LoRA). SPP injects sensor-aware conditioning into the network via learnable prompts derived from RAW pixel distribution priors, while LoRA enables efficient task-specific tuning by updating only low-rank matrices in key backbone layers. Despite minimal tuning, Dr. RAW delivers superior results across four RAW-based tasks (object detection, semantic segmentation, instance segmentation, and pose estimation) on nine datasets encompassing various light conditions. By harnessing the intrinsic physical cues of RAW alongside parameter-efficient techniques, Dr. RAW advances RAW-based vision systems, achieving both high accuracy and computational economy. The source code is available here.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e46b89e5-12f5-48e4-bbfa-959fb997c590Builds on24
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Segment Anything in High QualityLei Ke, Mingqiao Ye, Martin Danelljan, Yifan Liu et al.NeurIPS 2023 · 709 citations
- Surgical Fine-Tuning Improves Adaptation to Distribution ShiftsYoonho Lee, Annie S. Chen, Fahim Tajwar, Ananya Kumar et al.ICLR 2023 · 47 citations
- ReconfigISP: Reconfigurable Camera Image Processing PipelineKe Yu, Zexian Li, Yue Peng, Chen Change Loy et al.ICCV 2021 · 46 citations
Related papers
- Beyond RGB: Adaptive Parallel Processing for RAW Object DetectionShani Gamrian, Hila Barel, Feiran Li, Masakazu Yoshimura et al.ICCV 2025 · 4 citations
- Task-Aware Image Signal Processor for Advanced Visual PerceptionKai Chen, Jin Xiao, Leheng Zhang, Kexuan Shi et al.CVPR 2026 · 3 citations
- SpiralDiff: Spiral Diffusion with LoRA for RGB-to-RAW Conversion Across CamerasHuanjing Yue, Shangbin Xie, Cong Cao, Qian Wu et al.CVPR 2026
- LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation modelsZiqi Lu, Heng Yang, Danfei Xu, Boyi Li et al.ICLR 2025
- Correlated Low-Rank Adaptation for ConvNetsWu Ran, Weijia Zhang, Shuyang Pang, Qi Zhu et al.NeurIPS 2025 · 5 citations
