SmartPortraits: Depth Powered Handheld Smartphone Dataset of Human Portraits for State Estimation, Reconstruction and Synthesis
Anastasiia Kornilova, Marsel Faizullin, Konstantin Pakulev, Andrey Sadkov, Denis Kukushkin, Azat Akhmetyanov, Timur Akhtyamov, Hekmat Taherinejad, Gonzalo Ferrer
摘要
We present a dataset of 1000 video sequences of human portraits recorded in real and uncontrolled conditions by using a handheld smartphone accompanied by an external high-quality depth camera. The collected dataset contains 200 people captured in different poses and locations and its main purpose is to bridge the gap between raw measurements obtained from a smartphone and downstream applications, such as state estimation, 3D reconstruction, view synthesis, etc. The sensors employed in data collection are the smartphone's camera and Inertial Measurement Unit (IMU), and an external Azure Kinect DK depth camera software synchronized with sub-millisecond precision to the smartphone system. During the recording, the smartphone flash is used to provide a periodic secondary source of lightning. Accurate mask of the foremost person is provided as well as its impact on the camera alignment accuracy. For evaluation purposes, we compare multiple state-of-the-art camera alignment methods by using a Motion Cap-ture system. We provide a smartphone visual-inertial bench-mark for portrait capturing, where we report results for multiple methods and motivate further use of the provided trajectories, available in the dataset, in view synthesis and 3D reconstruction tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- STream3R: Scalable Sequential 3D Reconstruction with Causal TransformerYushi Lan, Yihang Luo, Fangzhou Hong, Shangchen Zhou 等ICLR 2026 · 被引用 84 次
- CorresNeRF: Image Correspondence Priors for Neural Radiance FieldsYixing Lao, Xiaogang Xu, Zhipeng Cai, Xihui Liu 等NeurIPS 2023 · 被引用 21 次
- Depth Pro: Sharp Monocular Metric Depth in Less Than a SecondAlexey Bochkovskiy, Amaël Delaunoy, Hugo Germain, Marcel Santos 等ICLR 2025 · 被引用 15 次
- NeuWigs: A Neural Dynamic Model for Volumetric Hair Capture and AnimationZiyan Wang, Giljoo Nam, Tuur Stuyck, Stephen Lombardi 等CVPR 2023
- LightIt: Illumination Modeling and Control for Diffusion ModelsPeter Kocsis, Julien Philip, Kalyan Sunkavalli, Matthias Nießner 等CVPR 2024
它引用的顶会 Paper14
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima 等ICCV 2019 · 被引用 1,411 次
- Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsEgor Zakharov, Aliaksandra Shysheya, Egor Burkov, Victor S. LempitskyICCV 2019 · 被引用 687 次
- Planar Prior Assisted PatchMatch Multi-View StereoQingshan Xu, Wenbing TaoAAAI 2020 · 被引用 154 次
- 3DPeople: Modeling the Geometry of Dressed HumansAlbert Pumarola, Jordi Sanchez, Gary P. T. Choi, Alberto Sanfeliu 等ICCV 2019 · 被引用 137 次
相关 Paper
- Pose-on-the-Go: Approximating User Pose with Smartphone Sensor Fusion and Inverse KinematicsKaran Ahuja, Sven Mayer, Mayank Goel, Chris HarrisonCHI 2021 · 被引用 40 次
- 100-Phones: A Large VI-SLAM Dataset for Augmented Reality Towards Mass Deployment on Mobile PhonesGuofeng Zhang, Jin Yuan, Haomin Liu, Zhen Peng 等IEEE VR 2024 · 被引用 5 次
- Robust Inertial Motion Tracking through Deep Sensor Fusion across Smart Earbuds and SmartphoneJian Gong, Xinyu Zhang, Yuanjun Huang, Ju Ren 等UbiComp 2021 · 被引用 37 次
- EMDB: The Electromagnetic Database of Global 3D Human Pose and Shape in the WildManuel Kaufmann, Jie Song, Chen Guo, Kaiyue Shen 等ICCV 2023 · 被引用 94 次
- Multi-Sensor Large-Scale Dataset for Multi-View 3D ReconstructionOleg Voynov, Gleb Bobrovskikh, Pavel A. Karpyshev, Saveliy Galochkin 等CVPR 2023
