OCTDiff: Bridged Diffusion Model for Portable OCT Super-Resolution and Enhancement
Ye Tian, Angela McCarthy, Gabriel Gomide, Nancy Liddle, Jedrzej Golebka, Royce W. S. Chen, Jeffrey M. Liebmann, Kaveri A. Thakoor
Abstract
Medical imaging super-resolution is critical for improving diagnostic utility and reducing costs, particularly for low-cost modalities such as portable Optical Co-herence Tomography (OCT). We propose OCTDiff, a bridged diffusion model designed to enhance image resolution and quality from portable OCT devices. Our image-to-image diffusion framework addresses key challenges in the conditional generation process of denoising diffusion probabilistic models (DDPMs). We introduce Adaptive Noise Aggregation (ANA), a novel module to improve denoising dynamics within the reverse diffusion process. Additionally, we integrate Multi-Scale Cross-Attention (MSCA) into the U-Net backbone to capture local dependencies across spatial resolutions. To address overfitting on small clinical datasets and to preserve fine structural details essential for retinal diagnostics, we design a customized loss function guided by clinical quality scores. OCTDiff out-performs convolutional baselines and standard DDPMs, achieving state-of-the-art performance on clinical portable OCT datasets. Our model and its downstream applications have the potential to generalize to other medical imaging modalities and revolutionize the current workflow of ophthalmic diagnostics. The code is available at https://github.com/AI4VSLab/OCTDiff.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- DCAFuse: Dual-Branch Diffusion-CNN Complementary Feature Aggregation Network for Multi-Modality Image FusionXudong Lu, Yuqi Jiang, Haiwen Hong, Qi Sun et al.ACM MM 2024 · 8 citations
- MicroDiffusion: Implicit Representation-Guided Diffusion for 3D Reconstruction from Limited 2D Microscopy ProjectionsMude Hui, Zihao Wei, Hongru Zhu, Fei Xia et al.CVPR 2024
- MedSegDiff-V2: Diffusion-Based Medical Image Segmentation with TransformerJunde Wu, Wei Ji, Huazhu Fu, Min Xu et al.AAAI 2024 · 311 citations
- Beyond Visual Quality: Fidelity-Oriented Diffusion Model for Real-world Image Super-ResolutionZhenxuan Fang, Shuaibo Wang, Weisheng Dong, Junwei Xu et al.ACM MM 2025
- A Unified Conditional Framework for Diffusion-based Image RestorationYi Zhang, Xiaoyu Shi, Dasong Li, Xiaogang Wang et al.NeurIPS 2023 · 44 citations
