Goal Conditioned Reinforcement Learning for Photo Finishing Tuning
Jiarui Wu, Yujin Wang, Lingen Li, Zhang Fan, Tianfan Xue
Abstract
Photo finishing tuning aims to automate the manual tuning process of the photo finishing pipeline, like Adobe Lightroom or Darktable. Previous works either use zeroth-order optimization, which is slow when the set of parameters increases, or rely on a differentiable proxy of the target finishing pipeline, which is hard to train. To overcome these challenges, we propose a novel goal-conditioned reinforcement learning framework for efficiently tuning parameters using a goal image as a condition. Unlike previous approaches, our tuning framework does not rely on any proxy and treats the photo finishing pipeline as a black box. Utilizing a trained reinforcement learning policy, it can efficiently find the desired set of parameters within just 10 queries, while optimization based approaches normally take 200 queries. Furthermore, our architecture utilizes a goal image to guide the iterative tuning of pipeline parameters, allowing for flexible conditioning on pixel-aligned target images, style images, or any other visually representable goals. We conduct detailed experiments on photo finishing tuning and photo stylization tuning tasks, demonstrating the advantages of our method. Project website: https://openimaginglab.github.io/RLPixTuner/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 975bce2d-abf8-423c-9021-ac126a8b2b10Cited by top-tier papers4
- JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching AgentYunlong Lin, Zixu Lin, Kunjie Lin, Jinbin Bai et al.NeurIPS 2025 · 42 citations
- Fine-grained Image Aesthetic Assessment: Learning Discriminative Scores from Relative RanksZhichao Yang, Jianjie Wang, Zhixianhe Zhang, Pangu Xie et al.CVPR 2026 · 5 citations
- Learning to Clean: Reinforcement Learning for Noisy Label CorrectionMarzi Heidari, Hanping Zhang, Yuhong GuoNeurIPS 2025
- InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral SpaceJiarui Wu, Yujin Wang, Ruikang Li, Fan Zhang et al.CVPR 2026
Builds on7
- Unpaired Image Enhancement Featuring Reinforcement-Learning-Controlled Image Editing SoftwareSatoshi Kosugi, Toshihiko YamasakiAAAI 2020 · 104 citations
- ReconfigISP: Reconfigurable Camera Image Processing PipelineKe Yu, Zexian Li, Yue Peng, Chen Change Loy et al.ICCV 2021 · 46 citations
- SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color EditingJing Shi, Ning Xu, Haitian Zheng, Alex Smith et al.CVPR 2022 · 15 citations
- InstructPix2Pix: Learning to Follow Image Editing InstructionsTim Brooks, Aleksander Holynski, Alexei A. EfrosCVPR 2023
- Neural Auto-Exposure for High-Dynamic Range Object DetectionEmmanuel Onzon, Fahim Mannan, Felix HeideCVPR 2021
Related papers
- AutoEdit: Automatic Hyperparameter Tuning for Image EditingChau Pham, Quan Dao, Mahesh Bhosale, Yunjie Tian et al.NeurIPS 2025 · 3 citations
- Optimizing Prompts for Text-to-Image GenerationYaru Hao, Zewen Chi, Li Dong, Furu WeiNeurIPS 2023 · 303 citations
- RL-SeqISP: Reinforcement Learning-Based Sequential Optimization for Image Signal ProcessingXinyu Sun, Zhikun Zhao, Lili Wei, Congyan Lang et al.AAAI 2024 · 13 citations
- Reinforcement Learning for Fine-tuning Text-to-Image Diffusion ModelsYing Fan, Olivia Watkins, Yuqing Du, Hao Liu et al.NeurIPS 2023 · 372 citations
- Directly Fine-Tuning Diffusion Models on Differentiable RewardsKevin Clark, Paul Vicol, Kevin Swersky, David J. FleetICLR 2024 · 377 citations
