View-LSTM: Novel-View Video Synthesis Through View Decomposition
Mohamed Ilyes Lakhal, Oswald Lanz, Andrea Cavallaro
摘要
We tackle the problem of synthesizing a video of multiple moving people as seen from a novel view, given only an input video and depth information or human poses of the novel view as prior. This problem requires a model that learns to transform input features into target features while maintaining temporal consistency. To this end, we learn an invariant feature from the input video that is shared across all viewpoints of the same scene and a view-dependent feature obtained using the target priors. The proposed approach, View-LSTM, is a recurrent neural network structure that accounts for the temporal consistency and target feature approximation constraints. We validate View-LSTM by designing an end-to-end generator for novel-view video synthesis. Experiments on a large multi-view action recognition dataset validate the proposed model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Rotationally-Consistent Novel View Synthesis for HumansYoungjoong Kwon, Stefano Petrangeli, Dahun Kim, Haoliang Wang 等ACM MM 2020 · 被引用 5 次
- Consistent depth of moving objects in videoZhoutong Zhang, Forrester Cole, Richard Tucker, William T. Freeman 等SIGGRAPH 2021 · 被引用 26 次
- Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single ImageXuanchi Ren, Xiaolong WangCVPR 2022 · 被引用 42 次
- MultiDiff: Consistent Novel View Synthesis from a Single ImageNorman Müller, Katja Schwarz, Barbara Rössle, Lorenzo Porzi 等CVPR 2024 · 被引用 14 次
- AR-1-to-3: Single Image to Consistent 3D Object via Next-View PredictionXuying Zhang, Yupeng Zhou, Kai Wang, Yikai Wang 等ICCV 2025 · 被引用 3 次
