Mask-ToF: Learning Microlens Masks for Flying Pixel Correction in Time-of-Flight Imaging
Ilya Chugunov, Seung-Hwan Baek, Qiang Fu, Wolfgang Heidrich, Felix Heide
摘要
We introduce Mask-ToF, a method to reduce flying pixels (FP) in time-of-flight (ToF) depth captures. FPs are pervasive artifacts which occur around depth edges, where light paths from both an object and its background are integrated over the aperture. This light mixes at a sensor pixel to produce erroneous depth estimates, which can adversely affect downstream 3D vision tasks. Mask-ToF starts at the source of these FPs, learning a microlens-level occlusion mask which effectively creates a custom-shaped sub-aperture for each sensor pixel. This modulates the selection of foreground and background light mixtures on a per-pixel basis and thereby encodes scene geometric information directly into the ToF measurements. We develop a differentiable ToF simulator to jointly train a convolutional neural network to decode this information and produce high-fidelity, low-FP depth reconstructions. We test the effectiveness of Mask-ToF on a simulated light field dataset and validate the method with an experimental prototype. To this end, we manufacture the learned amplitude mask and design an optical relay system to virtually place it on a high-resolution ToF sensor. We find that Mask-ToF generalizes well to real data without retraining, cutting FP counts in half.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Seeing through obstructions with diffractive cloakingZheng Shi, Yuval Bahat, Seung-Hwan Baek, Qiang Fu 等SIGGRAPH 2022 · 被引用 34 次
- Fisher Information Guidance for Learned Time-of-Flight ImagingJiaqu Li, Tao Yue, Sijie Zhao, Xuemei HuCVPR 2022 · 被引用 7 次
- Collaborative On-Sensor Array CamerasJipeng Sun, Kaixuan Wei, Thomas Eboli, Congli Wang 等SIGGRAPH 2025 · 被引用 2 次
- LocIn: Inferring Semantic Location from Spatial Maps in Mixed RealityHabiba Farrukh, Reham Mohamed, Aniket Nare, Antonio Bianchi 等USENIX Security 2023
- Latent Space ImagingMatheus Souza, Yidan Zheng, Kaizhang Kang, Yogeshwar Nath Mishra 等CVPR 2025
它引用的顶会 Paper6
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess 等ICCV 2019 · 被引用 2,966 次
- Depth From Videos in the Wild: Unsupervised Monocular Depth Learning From Unknown CamerasAriel Gordon, Hanhan Li, Rico Jonschkowski, Anelia AngelovaICCV 2019 · 被引用 397 次
- Deep Optics for Monocular Depth Estimation and 3D Object DetectionJulie Chang, Gordon WetzsteinICCV 2019 · 被引用 219 次
- AANet: Adaptive Aggregation Network for Efficient Stereo MatchingHaofei Xu, Juyong ZhangCVPR 2020
- Single-Shot Monocular RGB-D Imaging Using Uneven Double RefractionAndreas Meuleman, Seung-Hwan Baek, Felix Heide, Min H. KimCVPR 2020
相关 Paper
- RADU: Ray-Aligned Depth Update Convolutions for ToF Data DenoisingMichael Schelling, Pedro Hermosilla, Timo RopinskiCVPR 2022 · 被引用 18 次
- Dense Metric Depth Completion from Sparse Direct Time-of-Flight SensorsHakyeong Kim, Ruicheng Wang, Chengtang Yao, Jiaolong Yang 等CVPR 2026 · 被引用 1 次
- InDepth: Real-time Depth Inpainting for Mobile Augmented RealityYunfan Zhang, Tim Scargill, Ashutosh Vaishnav, Gopika Premsankar 等UbiComp 2022 · 被引用 27 次
- Non-line-of-sight imaging with arbitrary relay surface geometries via 3D Gaussian Transient RenderingYi Wang, Ziyu Zhan, Yuran Wang, Hao Wang 等SIGGRAPH 2026
- TöRF: Time-of-Flight Radiance Fields for Dynamic Scene View SynthesisBenjamin Attal, Eliot Laidlaw, Aaron Gokaslan, Changil Kim 等NeurIPS 2021 · 被引用 140 次
