Hashing Neural Video Decomposition with Multiplicative Residuals in Space-Time
Cheng-Hung Chan, Cheng-Yang Yuan, Cheng Sun, Hwann-Tzong Chen
摘要
We present a video decomposition method that facilitates layer-based editing of videos with spatiotemporally varying lighting and motion effects. Our neural model decomposes an input video into multiple layered representations, each comprising a 2D texture map, a mask for the original video, and a multiplicative residual characterizing the spatiotemporal variations in lighting conditions. A single edit on the texture maps can be propagated to the corresponding locations in the entire video frames while preserving other contents’ consistencies. Our method efficiently learns the layer-based neural representations of a 1080p video in 25s per frame via coordinate hashing and allows real-time rendering of the edited result at 71 fps on a single GPU. Qualitatively, we run our method on various videos to show its effectiveness in generating high-quality editing effects. Quantitatively, we propose to adopt feature-tracking evaluation metrics for objectively assessing the consistency of video editing. Project page: https://lightbulb12294.github.io/hashing-nvd/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Splatter a Video: Video Gaussian Representation for Versatile ProcessingYang-Tian Sun, Yihua Huang, Lin Ma, Xiaoyang Lyu 等NeurIPS 2024 · 被引用 41 次
- NaRCan: Natural Refined Canonical Image with Integration of Diffusion Prior for Video EditingTing-Hsuan Chen, Jiewen Chan, Hau-Shiang Shiu, Shih-Han Yen 等NeurIPS 2024 · 被引用 8 次
- HyperNVD: Accelerating Neural Video Decomposition via HypernetworksMaria Pilligua, Danna Xue, Javier Vazquez-CorralCVPR 2025
它引用的顶会 Paper10
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Space-Time Correspondence as a Contrastive Random WalkAllan Jabri, Andrew Owens, Alexei A. EfrosNeurIPS 2020 · 被引用 356 次
- GAN-Supervised Dense Visual AlignmentWilliam S. Peebles, Jun-Yan Zhu, Richard Zhang, Antonio Torralba 等CVPR 2022 · 被引用 50 次
- Deformable Sprites for Unsupervised Video DecompositionVickie Ye, Zhengqi Li, Richard Tucker, Angjoo Kanazawa 等CVPR 2022 · 被引用 45 次
- Learning Pixel Trajectories with Multiscale Contrastive Random WalksZhangxing Bian, Allan Jabri, Alexei A. Efros, Andrew OwensCVPR 2022 · 被引用 35 次
相关 Paper
- Video Decomposition Prior: Editing Videos Layer by LayerGaurav Shrivastava, Ser-Nam Lim, Abhinav ShrivastavaICLR 2024 · 被引用 11 次
- Editable free-viewpoint video using a layered neural representationJiakai Zhang, Xinhang Liu, Xinyi Ye, Fuqiang Zhao 等SIGGRAPH 2021 · 被引用 80 次
- RT-VENet: A Convolutional Network for Real-time Video EnhancementMohan Zhang, Qiqi Gao, Jinglu Wang, Henrik Turbell 等ACM MM 2020 · 被引用 5 次
- STRIVE: Scene Text Replacement In VideosVijay Kumar B. G, Jeyasri Subramanian, Varnith Chordia, Eugene Bart 等ICCV 2021 · 被引用 14 次
- StableVideo: Text-driven Consistency-aware Diffusion Video EditingWenhao Chai, Xun Guo, Gaoang Wang, Yan LuICCV 2023 · 被引用 219 次
