Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
Yang Yang, Siming Zheng, Qirui Yang, Jinwei Chen, Boxi Wu, Xiaofei He, Deng Cai, Bo Li, Peng-Tao Jiang
摘要
Diffusion models have recently emerged as powerful tools for camera simulation, enabling both geometric transformations and realistic optical effects. Among these, image-based bokeh rendering has shown promising results, but diffusion for video bokeh remains unexplored. Existing image-based methods are plagued by temporal flickering and inconsistent blur transitions, while current video editing methods lack explicit control over the focus plane and bokeh intensity. These issues limit their applicability for controllable video bokeh. In this work, we propose a one-step diffusion framework for generating temporally coherent, depth-aware video bokeh rendering. The framework employs a multi-plane image (MPI) representation adapted to the focal plane to condition the video diffusion model, thereby enabling it to exploit strong 3D priors from pretrained backbones. To further enhance temporal stability, depth robustness, and detail preservation, we introduce a progressive training strategy. Experiments on synthetic and real-world benchmarks demonstrate superior temporal coherence, spatial accuracy, and controllability, outperforming prior baselines. This work represents the first dedicated diffusion framework for video bokeh generation, establishing a new baseline for temporally coherent and controllable depth-of-field effects. Project page is available at this website https://vivocameraresearch.github.io/any2bokeh/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Elastic3D: Controllable Stereo Video Conversion with Guided Latent DecodingNando Metzger, Prune Truong, Goutam Bhat, Konrad Schindler 等CVPR 2026 · 被引用 3 次
- Preserving Source Video Realism: High-Fidelity Face Swapping for Cinematic QualityZekai Luo, Zongze Du, Zhouhang Zhu, Hao Zhong 等CVPR 2026 · 被引用 1 次
- Towards Photorealistic and Efficient Bokeh Rendering via Diffusion FrameworkLinxiao Shi, Siming Zheng, Zerong Wang, Hao Zhang 等CVPR 2026
它引用的顶会 Paper14
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific TuningYuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang 等ICLR 2024 · 被引用 1,493 次
- DragDiffusion: Harnessing Diffusion Models for Interactive Point-Based Image EditingYujun Shi, Chuhui Xue, Jun Hao Liew, Jiachun Pan 等CVPR 2024 · 被引用 117 次
- BokehMe: When Neural Rendering Meets Classical RenderingJuewen Peng, Zhiguo Cao, Xianrui Luo, Hao Lu 等CVPR 2022 · 被引用 45 次
- Dr.Bokeh: DiffeRentiable Occlusion-Aware Bokeh RenderingYichen Sheng, Zixun Yu, Lu Ling, Zhiwen Cao 等CVPR 2024 · 被引用 9 次
相关 Paper
- BokehCrafter: Taming Video Diffusion Models for Controllable Bokeh RenderingQiwen Wang, Liao Shen, Jiaqi Li, Tianqi Liu 等AAAI 2026
- Video Bokeh Rendering: Make Casual Videography CinematicYawen Luo, Min Shi, Liao Shen, Yachuan Huang 等ACM MM 2024 · 被引用 1 次
- BokehFlow: Depth-Free Controllable Bokeh Rendering via Flow MatchingYachuan Huang, Xianrui Luo, Qiwen Wang, Liao Shen 等AAAI 2026 · 被引用 2 次
- UniScene-MoTion: Unified Scene & Motion-aware Diffusion Transition FrameworkRui Jiang, Chongmian Wang, Xinghe Fu, Yehao Lu 等AAAI 2026
- BokehDiff: Neural Lens Blur with One-Step DiffusionChengxuan Zhu, Qingnan Fan, Qi Zhang, Jinwei Chen 等ICCV 2025 · 被引用 2 次
