Stabilizing and Accelerating Autofocus with Expert Trajectory Regularized Deep Reinforcement Learning
Shouhang Zhu, Chenglin Li, Yuankun Jiang, Li Wei, Nuowen Kan, Ziyang Zheng, Wenrui Dai, Junni Zou, Hongkai Xiong
Abstract
Autofocus is a crucial component of modern digital cameras. While recent learning-based methods achieve stateof-the-art in focus prediction accuracy, they unfortunately ignore the potential focus hunting phenomenon of backand-forth lens movement in the multi-step focusing procedure. To address this, in this paper, we propose an expert regularized deep reinforcement learning (DRL)-based approach for autofocus, which can utilize the sequential information of lens movement trajectory to both enhance the multi-step in-focus prediction accuracy and reduce the chance of focus hunting. Our method generally follows an actor-critic framework. To accelerate the DRL's training with a higher sample efficiency, we initialize the policy with a pre-trained single-step prediction network, where the network is further improved by modifying the output of absolute in-focus position distribution to the relative lens movement distribution to establish a better mapping between input images and lens movement. To further stabilize DRL's training with a lower occurrence of focus hunting in the resulting lens movement trajectory, we generate some offline trajectories based on prior knowledge to avoid focus hunting, which are then leveraged as an offline dataset of expert trajectories to regularize the actor network's training. Empirical evaluations show that our method outperforms those learning-based methods on public benchmarks, with higher single-and multi-step prediction accuracy, and a significant reduction of focus hunting rate.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 034de11f-a402-46cb-83b1-b78fff050e7dBuilds on12
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu et al.CVPR 2024 · 847 citations
- Learning Single Camera Depth Estimation Using Dual-PixelsRahul Garg, Neal Wadhwa, Sameer Ansari, Jonathan T. BarronICCV 2019 · 123 citations
- Accelerating Robotic Reinforcement Learning via Parameterized Action PrimitivesMurtaza Dalal, Deepak Pathak, Ruslan SalakhutdinovNeurIPS 2021 · 121 citations
- TorchRL: A data-driven decision-making library for PyTorchAlbert Bou, Matteo Bettini, Sebastian Dittert, Vikash Kumar et al.ICLR 2024 · 77 citations
- Learning to Reduce Defocus Blur by Realistically Modeling Dual-Pixel DataAbdullah Abuolaim, Mauricio Delbracio, Damien Kelly, Michael S. Brown et al.ICCV 2021 · 72 citations
Related papers
- Learning to AutofocusCharles Herrmann, Richard Strong Bowen, Neal Wadhwa, Rahul Garg et al.CVPR 2020
- Optical Model-Driven Sharpness Mapping for Autofocus in Small Depth-of-Field and Severe Defocus ScenariosChen-Liang Fan, Mingpei Cao, Chih Chien Hung, Yuesheng ZhuICCV 2025 · 1 citation
- Exploring Positional Characteristics of Dual-Pixel Data for Camera AutofocusMyungsub Choi, Hana Lee, Hyong-Euk LeeICCV 2023 · 9 citations
- Cost-Aware Fine-Grained Recognition for IoTs Based on Sequential FixationsHanxiao Wang, Venkatesh Saligrama, Stan Sclaroff, Vitaly AblavskyICCV 2019 · 2 citations
- Reinforcement Learning with Adaptive Regularization for Safe Control of Critical SystemsHaozhe Tian, Homayoun Hamedmoghadam, Robert Shorten, Pietro FerraroNeurIPS 2024 · 11 citations
