User-Instructed Disparity-aware Defocus Control
Yudong Han, Yan Yang, Hao Yang, Liyuan Pan
摘要
In photography, an All-in-Focus (AiF) image may not always effectively convey the creator’s intent. Professional photographers manipulate Depth of Field (DoF) to control which regions appear sharp or blurred, achieving compelling artistic effects. For general users, the ability to flexibly adjust DoF enhances creative expression and image quality. In this paper, we propose UiD, a User-Instructed DoF control framework, that allows users to specify refocusing regions using text, box, or point prompts, and our UiD automatically simulates in-focus and out-of-focus (OoF) regions in the given images. However, controlling defocus blur in a single-lens camera remains challenging due to the difficulty in estimating depth-aware aberrations and the suboptimal quality of reconstructed AiF images. To address this, we leverage dual-pixel (DP) sensors, commonly found in DSLR-style and mobile cameras. DP sensors provide a small-baseline stereo pair in a single snapshot, enabling depth-aware aberration estimation. Our approach first establishes an invertible mapping between OoF and AiF images to learn spatially varying defocus kernels and the disparity features. These depth-aware kernels enable bidirectional image transformation—deblurring out-of-focus (OoF) images into all-in-focus (AiF) representations, and conversely reblurring AiF images into OoF outputs—by seamlessly switching between the kernel and its inverse form. These depth-aware kernels enable both deblurring of OoF images into AiF representations and reblurring AiF images into OoF representations by flexibly switching its original form to its inverse one. For user-guided refocusing, we first generate masks based on user prompts using SAM, which modulates disparity features in closed form, allowing dynamic kernel re-estimation for reblurring. This achieves user-controlled refocusing effects. Extensive experiments on both common datasets and the self-collected dataset demonstrate that UiD offers superior flexibility and quality in DoF manipulation imaging.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- Omni-Kernel Network for Image RestorationYuning Cui, Wenqi Ren, Alois KnollAAAI 2024 · 被引用 290 次
- Single Image Defocus Deblurring Using Kernel-Sharing Parallel Atrous ConvolutionsHyeongseok Son, Junyong Lee, Sunghyun Cho, Seungyong LeeICCV 2021 · 被引用 124 次
相关 Paper
- Defocus Map Estimation and Deblurring from a Single Dual-Pixel ImageShumian Xin, Neal Wadhwa, Tianfan Xue, Jonathan T. Barron 等ICCV 2021 · 被引用 47 次
- DC2: Dual-Camera Defocus Control by Learning to RefocusHadi Alzayer, Abdullah Abuolaim, Leung Chun Chan, Yang Yang 等CVPR 2023
- Dual Pixel Exploration: Simultaneous Depth Estimation and Image RestorationLiyuan Pan, Shah Chowdhury, Richard Hartley, Miaomiao Liu 等CVPR 2021
- Learnable Blur Kernel for Single-Image Defocus Deblurring in the WildJucai Zhai, Pengcheng Zeng, Chihao Ma, Jie Chen 等AAAI 2023 · 被引用 7 次
- K3DN: Disparity-Aware Kernel Estimation for Dual-Pixel Defocus DeblurringYan Yang, Liyuan Pan, Liu Liu, Miaomiao LiuCVPR 2023
