Consistent Direct Time-of-Flight Video Depth Super-Resolution
Zhanghao Sun, Wei Ye, Jinhui Xiong, Gyeongmin Choe, Jialiang Wang, Shuochen Su, Rakesh Ranjan
Abstract
Figure 1. We propose the first multi-frame approaches, dToF depth video super-resolution (DVSR) and histogram video super-resolution (HVSR), to super-resolve low-resolution dToF sensor videos with the high-resolution RGB frame guidance. The point cloud visualizations of depth predictions reveal that, by utilizing multi-frame correlation, DVSR predicts significantly better geometry compared to state-of-theart per-frame depth enhancement networks [42] while being more lightweight; HVSR further improves the fidelity of geometry and reduces flying pixels by utilizing the dToF histogram information. Besides the improvements in per-frame estimation, we highly recommend readers to check out the supplementary video, which visualizes the significant improvements in temporal stability across the entire sequences.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8f5f417e-bab3-4512-a35d-40628aff4078Cited by top-tier papers9
- Self-Distilled Depth Refinement with Noisy Poisson FusionJiaqi Li, Yiran Wang, Jinghong Zheng, Zihao Huang et al.NeurIPS 2024 · 8 citations
- DEPTHOR: Depth Enhancement from a Practical Light-Weight dToF Sensor and RGB ImageJijun Xiang, Xuan Zhu, Xianqi Wang, Yu Wang et al.ICCV 2025 · 3 citations
- SpatioTemporal Difference Network for Video Depth Super-ResolutionZhengxue Wang, Yuan Wu, Xiang Li, Zhiqiang Yan et al.AAAI 2026 · 2 citations
- Consistent Time-of-Flight Depth Denoising via Graph-Informed Geometric AttentionWeida Wang, Changyong He, Jin Zeng, Di QiuICCV 2025 · 1 citation
- Video Depth without Video ModelsBingxin Ke, Dominik Narnhofer, Shengyu Huang, Lei Ke et al.CVPR 2025
Builds on17
- Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene UnderstandingMike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar et al.ICCV 2021 · 633 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Dynamic View Synthesis from Dynamic Monocular VideoChen Gao, Ayush Saraf, Johannes Kopf, Jia-Bin HuangICCV 2021 · 522 citations
- Consistent video depth estimationXuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen et al.SIGGRAPH 2020 · 321 citations
- NerfingMVS: Guided Optimization of Neural Radiance Fields for Indoor Multi-view StereoYi Wei, Shaohui Liu, Yongming Rao, Wang Zhao et al.ICCV 2021 · 286 citations
Related papers
- SVDC: Consistent Direct Time-of-Flight Video Depth Completion with Frequency Selective FusionXuan Zhu, Jijun Xiang, Xianqi Wang, Longliang Liu et al.CVPR 2025
- Zooming Slow-Mo: Fast and Accurate One-Stage Space-Time Video Super-ResolutionXiaoyu Xiang, Yapeng Tian, Yulun Zhang, Yun Fu et al.CVPR 2020
- Stereo Video Super-Resolution via Exploiting View-Temporal CorrelationsRuikang Xu, Zeyu Xiao, Mingde Yao, Yueyi Zhang et al.ACM MM 2021 · 20 citations
- Cross-view Resolution and Frame Rate Joint Enhancement for Binocular VideoPanda Pan, Yang Zhao, Yuan Chen, Wei Jia et al.ACM MM 2023 · 1 citation
- BridgeNet: A Joint Learning Network of Depth Map Super-Resolution and Monocular Depth EstimationQi Tang, Runmin Cong, Ronghui Sheng, Lingzhi He et al.ACM MM 2021 · 47 citations
