Self-Supervised Human Depth Estimation From Monocular Videos
Feitong Tan, Hao Zhu, Zhaopeng Cui, Siyu Zhu, Marc Pollefeys, Ping Tan
摘要
Previous methods on estimating detailed human depth often require supervised training with 'ground truth' depth data. This paper presents a self-supervised method that can be trained on YouTube videos without known depth, which makes training data collection simple and improves the generalization of the learned network. The self-supervised learning is achieved by minimizing a photo-consistency loss, which is evaluated between a video frame and its neighboring frames warped according to the estimated depth and the 3D non-rigid motion of the human body. To solve this non-rigid motion, we first estimate a rough SMPL model at each video frame and compute the non-rigid body motion accordingly, which enables self-supervised learning on estimating the shape details. Experiments demonstrate that our method enjoys better generalization and performs much better on data in the wild.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Registering Explicit to Implicit: Towards High-Fidelity Garment mesh Reconstruction from Single ImagesHeming Zhu, Lingteng Qiu, Yuda Qiu, Xiaoguang HanCVPR 2022 · 被引用 32 次
- RAFaRe: Learning Robust and Accurate Non-parametric 3D Face Reconstruction from Pseudo 2D&3D PairsLongwei Guo, Hao Zhu, Yuanxun Lu, Menghua Wu 等AAAI 2023 · 被引用 14 次
- Fusing the Old with the New: Learning Relative Camera Pose with Geometry-Guided UncertaintyBingbing Zhuang, Manmohan ChandrakerCVPR 2021
- Learning High Fidelity Depths of Dressed Humans by Watching Social Media Dance VideosYasamin Jafarian, Hyun Soo ParkCVPR 2021
- Recurrent Multi-View Alignment Network for Unsupervised Surface RegistrationWanquan Feng, Juyong Zhang, Hongrui Cai, Haofei Xu 等CVPR 2021
它引用的顶会 Paper6
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima 等ICCV 2019 · 被引用 1,411 次
- Multi-Garment Net: Learning to Dress 3D People From ImagesBharat Lal Bhatnagar, Garvita Tiwari, Christian Theobalt, Gerard Pons-MollICCV 2019 · 被引用 447 次
- DeepHuman: 3D Human Reconstruction From a Single ImageZerong Zheng, Tao Yu, Yixuan Wei, Qionghai Dai 等ICCV 2019 · 被引用 367 次
- Tex2Shape: Detailed Full Human Body Geometry From a Single ImageThiemo Alldieck, Gerard Pons-Moll, Christian Theobalt, Marcus A. MagnorICCV 2019 · 被引用 343 次
- TexturePose: Supervising Human Mesh Estimation With Texture ConsistencyGeorgios Pavlakos, Nikos Kolotouros, Kostas DaniilidisICCV 2019 · 被引用 109 次
相关 Paper
- Self-Supervised 3D Human Mesh Recovery from a Single Image with Uncertainty-Aware LearningGuoli Yan, Zichun Zhong, Jing HuaAAAI 2024 · 被引用 1 次
- Mining Supervision for Dynamic Regions in Self-Supervised Monocular Depth EstimationHoang Chuong Nguyen, Tianyu Wang, José M. Álvarez, Miaomiao LiuCVPR 2024 · 被引用 5 次
- iVS-Net: Learning Human View Synthesis from Internet VideosJunting Dong, Qi Fang, Tianshuo Yang, Qing Shuai 等ICCV 2023 · 被引用 9 次
- AltNeRF: Learning Robust Neural Radiance Field via Alternating Depth-Pose OptimizationKun Wang, Zhiqiang Yan, Huang Tian, Zhenyu Zhang 等AAAI 2024 · 被引用 6 次
- Self-Supervised Geometric Correspondence for Category-Level 6D Object Pose Estimation in the WildKaifeng Zhang, Yang Fu, Shubhankar Borse, Hong Cai 等ICLR 2023 · 被引用 8 次
