3D Multi-frame Fusion for Video Stabilization
Zhan Peng, Xinyi Ye, Weiyue Zhao, Tianqi Liu, Huiqiang Sun, Baopu Li, Zhiguo Cao
Abstract
In this paper, we present RStab, a novel framework for video stabilization that integrates 3D multi-frame fusion through volume rendering. Departing from conventional methods, we introduce a 3D multi-frame perspective to generate stabilized images, addressing the challenge of full-frame generation while preserving structure. The core of our RStab framework lies in Stabilized Rendering (SR), a volume rendering module, fusing multi-frame information in 3D space. Specifically, SR involves warping features and colors from multiple frames by projection, fusing them into descriptors to render the stabilized image. However, the precision of warped information depends on the projection accuracy, a factor significantly influenced by dynamic regions. In response, we introduce the Adaptive Ray Range (ARR) module to integrate depth priors, adaptively defining the sampling range for the projection process. Additionally, we propose Color Correction (CC) assisting geometric constraints with optical flow for accurate color aggregation. Thanks to the three modules, our RStab demonstrates superior performance compared with previous stabilizers in the field of view (FOV), image quality, and video stability across various datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext de7804fa-81c2-4e47-a396-a964573d79deCited by top-tier papers2
- GaVS: 3D-Grounded Video Stabilization via Temporally-Consistent Local Reconstruction and RenderingZinuo You, Stamatios Georgoulis, Anpei Chen, Siyu Tang et al.SIGGRAPH 2025 · 3 citations
- No Labels, No Look-Ahead: Unsupervised Online Video Stabilization with Classical PriorsKan Ren, Gang Wan, TAO LIUCVPR 2026 · 2 citations
Builds on16
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang et al.ICCV 2021 · 1,024 citations
- Point-NeRF: Point-based Neural Radiance FieldsQiangeng Xu, Zexiang Xu, Julien Philip, Sai Bi et al.CVPR 2022 · 510 citations
Related papers
- Hybrid Neural Fusion for Full-frame Video StabilizationYu-Lun Liu, Wei-Sheng Lai, Ming-Hsuan Yang, Yung-Yu Chuang et al.ICCV 2021 · 58 citations
- 3D Video Stabilization With Depth Estimation by CNN-Based OptimizationYao-Chih Lee, Kuan-Wei Tseng, Yu-Ta Chen, Chien-Cheng Chen et al.CVPR 2021
- Minimum Latency Deep Online Video StabilizationZhuofan Zhang, Zhen Liu, Ping Tan, Bing Zeng et al.ICCV 2023 · 27 citations
- Learning Video Stabilization Using Optical FlowJiyang Yu, Ravi RamamoorthiCVPR 2020
- DynIBaR: Neural Dynamic Image-Based RenderingZhengqi Li, Qianqian Wang, Forrester Cole, Richard Tucker et al.CVPR 2023
