Boosting Monocular Metric Depth Estimation via Bokeh Rendering
Hangwei Zhang, Armando Fortes, Tianyi Wei, Xingang Pan
摘要
Bokeh rendering and depth estimation share a fundamental optical connection, yet existing methods fail to fully exploit this reciprocity. Conventional bokeh pipelines rely heavily on noisy depth maps that inevitably introduce visual artifacts. Conversely, existing monocular depth models typically follow two flawed paradigms. Generative diffusion-based frameworks often lack consistent metric scale. Meanwhile, feed-forward metric depth models frequently fail in textureless or distant regions where defocus blur can provide geometric information. We propose BokehDepth, a two-stage framework that treats synthetic defocus as a supervision-free geometric signal. In the first stage, a physically grounded generative model produces calibrated bokeh stacks from a single sharp input without requiring prior depth input. Subsequently, a lightweight defocus-aware aggregation module integrates these stacks into the encoder of a depth estimation framework. This mechanism allows the model to extract consistent geometric features from the defocus dimension while keeping the decoder architecture unchanged. Experiments demonstrate that BokehDepth achieves superior visual bokeh fidelity compared to depth-dependent rendering baselines and consistently enhances the metric accuracy of state-of-the-art monocular depth models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper37
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
相关 Paper
- BokehDiff: Neural Lens Blur with One-Step DiffusionChengxuan Zhu, Qingnan Fan, Qi Zhang, Jinwei Chen 等ICCV 2025 · 被引用 2 次
- BokehFlow: Depth-Free Controllable Bokeh Rendering via Flow MatchingYachuan Huang, Xianrui Luo, Qiwen Wang, Liao Shen 等AAAI 2026 · 被引用 2 次
- Dr.Bokeh: DiffeRentiable Occlusion-Aware Bokeh RenderingYichen Sheng, Zixun Yu, Lu Ling, Zhiwen Cao 等CVPR 2024 · 被引用 9 次
- Repurposing Marigold for Zero-Shot Metric Depth Estimation via Defocus Blur CuesChinmay Talegaonkar, Nikhil Gandudi Suresh, Zachary Novack, Yash Belhe 等NeurIPS 2025 · 被引用 4 次
- Towards Photorealistic and Efficient Bokeh Rendering via Diffusion FrameworkLinxiao Shi, Siming Zheng, Zerong Wang, Hao Zhang 等CVPR 2026
