Self-Distilled Depth Refinement with Noisy Poisson Fusion
Jiaqi Li, Yiran Wang, Jinghong Zheng, Zihao Huang, Ke Xian, Zhiguo Cao, Jianming Zhang
Abstract
Depth refinement aims to infer high-resolution depth with fine-grained edges and details, refining low-resolution results of depth estimation models. The prevailing methods adopt tile-based manners by merging numerous patches, which lacks efficiency and produces inconsistency. Besides, prior arts suffer from fuzzy depth boundaries and limited generalizability. Analyzing the fundamental reasons for these limitations, we model depth refinement as a noisy Poisson fusion problem with local inconsistency and edge deformation noises. We propose the Self-distilled Depth Refinement (SDDR) framework to enforce robustness against the noises, which mainly consists of depth edge representation and edge-based guidance. With noisy depth predictions as input, SDDR generates low-noise depth edge representations as pseudo-labels by coarse-to-fine self-distillation. Edge-based guidance with edge-guided gradient loss and edge-based fusion loss serves as the optimization objective equivalent to Poisson fusion. When depth maps are better refined, the labels also become more noise-free. Our model can acquire strong robustness to the noises, achieving significant improvements in accuracy, edge quality, efficiency, and generalizability on five different benchmarks. Moreover, directly training another model with edge labels produced by SDDR brings improvements, suggesting that our method could help with training robust refinement models in future works.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6765e5a8-fbcb-4495-8cd0-af03a51fa2dcCited by top-tier papers6
- Any Resolution Any Geometry: From Multi-View To Multi-PatchWenqing Cui, Zhenyu Li, Mykola Lavreniuk, Jian Shi et al.CVPR 2026 · 2 citations
- One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution ImagesByeongjun Kwon, Munchurl KimICCV 2025 · 1 citation
- Hyden: A Hybrid Dual-Path Encoder for Monocular Geometry of High-resolution ImagesZaiwei Zhang, Marc Mapeke, Wei Ye, Rakesh Ranjan et al.ICLR 2026
- Spectral-Geometric Neural Fields for Pose-Free LiDAR View SynthesisYinuo Jiang, Jun Cheng, Yiran Wang, Cheng ChengCVPR 2026
- TacoDepth: Towards Efficient Radar-Camera Depth Estimation with One-stage FusionYiran Wang, Jiaqi Li, Chaoyi Hong, Ruibo Li et al.CVPR 2025
Builds on20
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene UnderstandingMike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar et al.ICCV 2021 · 633 citations
Related papers
- 3D Distillation: Improving Self-Supervised Monocular Depth Estimation on Reflective SurfacesXuepeng Shi, Georgi Dikov, Gerhard Reitmayr, Tae-Kyun Kim et al.ICCV 2023 · 10 citations
- Multi-Resolution Monocular Depth Map Fusion by Self-Supervised Gradient-Based CompositionYaqiao Dai, Renjiao Yi, Chenyang Zhu, Hongjun He et al.AAAI 2023 · 8 citations
- Exploiting Pseudo Labels in a Self-Supervised Learning Framework for Improved Monocular Depth EstimationAndra Petrovai, Sergiu NedevschiCVPR 2022 · 56 citations
- PatchFusion: An End-to-End Tile-Based Framework for High-Resolution Monocular Metric Depth EstimationZhenyu Li, Shariq Farooq Bhat, Peter WonkaCVPR 2024
- Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth EstimationXinhao Cai, Gensheng Pei, Zeren Sun, Yazhou Yao et al.CVPR 2026 · 2 citations
