Deep End-to-End Alignment and Refinement for Time-of-Flight RGB-D Module
Di Qiu, Jiahao Pang, Wenxiu Sun, Chengxi Yang
Abstract
Recently, it is increasingly popular to equip mobile RGB cameras with Time-of-Flight (ToF) sensors for active depth sensing. However, for off-the-shelf ToF sensors, one must tackle two problems in order to obtain high-quality depth with respect to the RGB camera, namely 1) online calibration and alignment; and 2) complicated error correction for ToF depth sensing. In this work, we propose a framework for jointly alignment and refinement via deep learning. First, a cross-modal optical flow between the RGB image and the ToF amplitude image is estimated for alignment. The aligned depth is then refined via an improved kernel predicting network that performs kernel normalization and applies the bias prior to the dynamic convolution. To enrich our data for end-to-end training, we have also synthesized a dataset using tools from computer graphics. Experimental results demonstrate the effectiveness of our approach, achieving state-of-the-art for ToF refinement. * Both authors contributed equally. Jiahao Pang is the corresponding author, this work was done while he was with SenseTime. (a) Unaligned erroneous depth image. (b) Our result.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74e4bd1e-4e6d-4926-b440-65d8834286d6Cited by top-tier papers6
- RADU: Ray-Aligned Depth Update Convolutions for ToF Data DenoisingMichael Schelling, Pedro Hermosilla, Timo RopinskiCVPR 2022 · 18 citations
- Robust 3D Object Detection Using Probabilistic Point Clouds From Single-Photon LidarsBhavya Goyal, Felipe Gutierrez-Barragan, Wei Lin, Andreas Velten et al.ICCV 2025 · 2 citations
- Learnable Fractional Reaction-Diffusion Dynamics for Under-Display ToF Imaging and BeyondXin Qiao, Matteo Poggi, Xing Wei, Pengchao Deng et al.ICCV 2025 · 2 citations
- Consistent Direct Time-of-Flight Video Depth Super-ResolutionZhanghao Sun, Wei Ye, Jinhui Xiong, Gyeongmin Choe et al.CVPR 2023
- Structure Aggregation for Cross-Spectral Stereo Image Guided DenoisingZehua Sheng, Zhu Yu, Xiongwei Liu, Si-Yuan Cao et al.CVPR 2023
Related papers
- InDepth: Real-time Depth Inpainting for Mobile Augmented RealityYunfan Zhang, Tim Scargill, Ashutosh Vaishnav, Gopika Premsankar et al.UbiComp 2022 · 27 citations
- TöRF: Time-of-Flight Radiance Fields for Dynamic Scene View SynthesisBenjamin Attal, Eliot Laidlaw, Aaron Gokaslan, Changil Kim et al.NeurIPS 2021 · 140 citations
- Consistent Time-of-Flight Depth Denoising via Graph-Informed Geometric AttentionWeida Wang, Changyong He, Jin Zeng, Di QiuICCV 2025 · 1 citation
- Mask-ToF: Learning Microlens Masks for Flying Pixel Correction in Time-of-Flight ImagingIlya Chugunov, Seung-Hwan Baek, Qiang Fu, Wolfgang Heidrich et al.CVPR 2021
- Multi-Modal Neural Radiance Field for Monocular Dense SLAM with a Light-Weight ToF SensorXinyang Liu, Yijin Li, Yanbin Teng, Hujun Bao et al.ICCV 2023 · 41 citations
