Shakes on a Plane: Unsupervised Depth Estimation from Unstabilized Photography
Ilya Chugunov, Yuxuan Zhang, Felix Heide
摘要
Modern mobile burst photography pipelines capture and merge a short sequence of frames to recover an enhanced image, but often disregard the 3D nature of the scene they capture, treating pixel motion between images as a 2D aggregation problem. We show that in a "long-burst", fortytwo 12-megapixel RAW frames captured in a two-second sequence, there is enough parallax information from natural hand tremor alone to recover high-quality scene depth. To this end, we devise a test-time optimization approach that fits a neural RGB-D representation to long-burst data and simultaneously estimates scene depth and camera motion. Our plane plus depth model is trained end-to-end, and performs coarse-to-fine refinement by controlling which multiresolution volume features the network has access to at what time during training. We validate the method experimentally, and demonstrate geometrically accurate depth reconstructions with no additional hardware or separate data pre-processing and pose-estimation steps.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- AONeuS: A Neural Rendering Framework for Acoustic-Optical Sensor FusionMohamad Qadri, Kevin Zhang, Akshay Hinduja, Michael Kaess 等SIGGRAPH 2024 · 被引用 23 次
- TurboSL: Dense, Accurate and Fast 3D by Neural Inverse Structured LightParsa Mirdehghan, Maxx Wu, Wenzheng Chen, David B. Lindell 等CVPR 2024 · 被引用 5 次
- Dark3R: Learning Structure from Motion in the DarkAndrew Y. Guo, Anagh Malik, SaiKiran Kumar Tedla, Yutong Dai 等CVPR 2026 · 被引用 3 次
- Neural Spline Fields for Burst Image Fusion and Layer SeparationIlya Chugunov, David Shustin, Ruyu Yan, Chenyang Lei 等CVPR 2024
- Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation ModelsYu Jiang, Hanwen Jiang, Ahmed Abdelkader, Wen-Sheng Chu 等CVPR 2026
它引用的顶会 Paper22
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- Block-NeRF: Scalable Large Scene Neural View SynthesisMatthew Tancik, Vincent Casser, Xinchen Yan, Sabeek Pradhan 等CVPR 2022 · 被引用 702 次
相关 Paper
- The Implicit Values of A Good Hand Shake: Handheld Multi-Frame Neural Depth RefinementIlya Chugunov, Yuxuan Zhang, Zhihao Xia, Xuaner Zhang 等CVPR 2022 · 被引用 11 次
- Consistent depth of moving objects in videoZhoutong Zhang, Forrester Cole, Richard Tucker, William T. Freeman 等SIGGRAPH 2021 · 被引用 26 次
- Mesoscopic Photogrammetry With an Unstabilized Phone CameraKevin C. Zhou, Colin L. V. Cooke, Jaehee Park, Ruobing Qian 等CVPR 2021
- Digital Gimbal: End-to-End Deep Image Stabilization With Learnable Exposure TimesOmer Dahary, Matan Jacoby, Alex M. BronsteinCVPR 2021
- Image as an Imu: Estimating Camera Motion From a Single Motion-Blurred ImageJerred Chen, Ronald ClarkICCV 2025
