DiffusionEdge: Diffusion Probabilistic Model for Crisp Edge Detection
Yunfan Ye, Kai Xu, Yuhang Huang, Renjiao Yi, Zhiping Cai
摘要
Limited by the encoder-decoder architecture, learning-based edge detectors usually have difficulty predicting edge maps that satisfy both correctness and crispness. With the recent success of the diffusion probabilistic model (DPM), we found it is especially suitable for accurate and crisp edge detection since the denoising process is directly applied to the original image size. Therefore, we propose the first diffusion model for the task of general edge detection, which we call Diffu-sionEdge. To avoid expensive computational resources while retaining the final performance, we apply DPM in the latent space and enable the classic cross-entropy loss which is uncertainty-aware in pixel level to directly optimize the parameters in latent space in a distillation manner. We also adopt a decoupled architecture to speed up the denoising process and propose a corresponding adaptive Fourier filter to adjust the latent features of specific frequencies. With all the technical designs, DiffusionEdge can be stably trained with limited resources, predicting crisp and accurate edge maps with much fewer augmentation strategies. Extensive experiments on four edge detection benchmarks demonstrate the superiority of DiffusionEdge both in correctness and crispness. On the NYUDv2 dataset, compared to the second best, we increase the ODS, OIS (without post-processing) and AC by 30.2%, 28.1% and 65.1%, respectively. Code: https://github.com/GuHuangAI/DiffusionEdge .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Hand1000: Generating Realistic Hands from Text with Only 1, 000 ImagesHaozhuo Zhang, Bin Zhu, Yu Cao, Yanbin HaoAAAI 2025 · 被引用 11 次
- SAUGE: Taming SAM for Uncertainty-Aligned Multi-Granularity Edge DetectionXing Liufu, Chaolei Tan, Xiaotong Lin, Yonggang Qi 等AAAI 2025 · 被引用 10 次
- HumanSAM: Classifying Human-Centric Forgery Videos in Human Spatial, Appearance, and Motion AnomalyChang Liu, Yunfan Ye, Fan Zhang, Qingyang Zhou 等ICCV 2025 · 被引用 6 次
- TRACE: Your Diffusion Model is Secretly an Instance Edge DetectorSanghyun Jo, Ziseok Lee, Wooyeol Lee, Jonghyun Choi 等ICLR 2026 · 被引用 4 次
- Environment-Agnostic Pose: Generating Environment-Independent Object Representations for 6D Pose EstimationShaobo Zhang, Yuhang Huang, Wanqing Zhao, Wei Zhao 等ICCV 2025 · 被引用 3 次
它引用的顶会 Paper12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
- Structured Denoising Diffusion Models in Discrete State-SpacesJacob Austin, Daniel D. Johnson, Jonathan Ho, Daniel Tarlow 等NeurIPS 2021 · 被引用 2,256 次
- Grad-TTS: A Diffusion Probabilistic Model for Text-to-SpeechVadim Popov, Ivan Vovk, Vladimir Gogoryan, Tasnima Sadekova 等ICML 2021 · 被引用 715 次
相关 Paper
- MEMO: Human-like Crisp Edge Detection Using Masked Edge PredictionJiaxin Cheng, Yue Wu, Yicong ZhouCVPR 2026 · 被引用 1 次
- Hierarchical Integration Diffusion Model for Realistic Image DeblurringZheng Chen, Yulun Zhang, Ding Liu, Bin Xia 等NeurIPS 2023 · 被引用 179 次
- The Treasure Beneath Multiple Annotations: An Uncertainty-Aware Edge DetectorCaixia Zhou, Yaping Huang, Mengyang Pu, Qingji Guan 等CVPR 2023
- DIRE for Diffusion-Generated Image DetectionZhendong Wang, Jianmin Bao, Wengang Zhou, Weilun Wang 等ICCV 2023 · 被引用 479 次
- Event-Diffusion: Event-Based Image Reconstruction and Restoration with Diffusion ModelsQuanmin Liang, Xiawu Zheng, Kai Huang, Yan Zhang 等ACM MM 2023 · 被引用 13 次
