PoseGen: Learning to Generate 3D Human Pose Dataset with NeRF
Mohsen Gholami, Rabab Ward, Z. Jane Wang
摘要
This paper proposes an end-to-end framework for generating 3D human pose datasets using Neural Radiance Fields (NeRF). Public datasets generally have limited diversity in terms of human poses and camera viewpoints, largely due to the resource-intensive nature of collecting 3D human pose data. As a result, pose estimators trained on public datasets significantly underperform when applied to unseen out-of-distribution samples. Previous works proposed augmenting public datasets by generating 2D-3D pose pairs or rendering a large amount of random data. Such approaches either overlook image rendering or result in suboptimal datasets for pre-trained models. Here we propose PoseGen, which learns to generate a dataset (human 3D poses and images) with a feedback loss from a given pre-trained pose estimator. In contrast to prior art, our generated data is optimized to improve the robustness of the pre-trained model. The objective of PoseGen is to learn a distribution of data that maximizes the prediction error of a given pre-trained model. As the learned data distribution contains OOD samples of the pre-trained model, sampling data from such a distribution for further fine-tuning a pre-trained model improves the generalizability of the model. This is the first work that proposes NeRFs for 3D human data generation. NeRFs are data-driven and do not require 3D scans of humans. Therefore, using NeRF for data generation is a new direction for convenient user-specific data generation. Our extensive experiments show that the proposed PoseGen improves two baseline models (SPIN and HybrIK) on four datasets with an average 6% relative improvement.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Spatial Reasoning with Vision-Language Models in Ego-Centric Multi-View ScenesMohsen Gholami, Ahmad Rezaei, Zhou Weimin, Sitong Mao 等ICLR 2026 · 被引用 67 次
- PoseSyn: Synthesizing Diverse 3D Pose Data from In-the-Wild 2D DataChangHee Yang, Hyeonseop Song, Seokhun Choi, Seungwoo Lee 等ICCV 2025 · 被引用 1 次
它引用的顶会 Paper18
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 被引用 1,139 次
- A-NeRF: Articulated Neural Radiance Fields for Learning Human Shape, Appearance, and PoseShih-Yang Su, Frank Yu, Michael Zollhöfer, Helge RhodinNeurIPS 2021 · 被引用 316 次
- DenseRaC: Joint 3D Pose and Shape Estimation by Dense Render-and-CompareYuanlu Xu, Song-Chun Zhu, Tony TungICCV 2019 · 被引用 204 次
- Putting People in their Place: Monocular Regression of 3D People in DepthYu Sun, Wu Liu, Qian Bao, Yili Fu 等CVPR 2022 · 被引用 152 次
- Moulding Humans: Non-Parametric 3D Human Shape Estimation From Single ImagesValentin Gabeur, Jean-Sébastien Franco, Xavier Martin, Cordelia Schmid 等ICCV 2019 · 被引用 140 次
相关 Paper
- SynBody: Synthetic Dataset with Layered Human Models for 3D Human Perception and ModelingZhitao Yang, Zhongang Cai, Haiyi Mei, Shuai Liu 等ICCV 2023 · 被引用 73 次
- GNeRF: GAN-based Neural Radiance Field without Posed CameraQuan Meng, Anpei Chen, Haimin Luo, Minye Wu 等ICCV 2021 · 被引用 222 次
- Shape, Pose, and Appearance from a Single Image via Bootstrapped Radiance Field InversionDario Pavllo, David Joseph Tan, Marie-Julie Rakotosaona, Federico TombariCVPR 2023
- Pose-Free Neural Radiance Fields via Implicit Pose RegularizationJiahui Zhang, Fangneng Zhan, Yingchen Yu, Kunhao Liu 等ICCV 2023 · 被引用 17 次
- 3D-aware Blending with Generative NeRFsHyunsu Kim, Gayoung Lee, Yunjey Choi, Jin-Hwa Kim 等ICCV 2023 · 被引用 14 次
