Rethinking Surgical Smoke: A Smoke-Type-Aware Laparoscopic Video Desmoking Method and Dataset
Qifan Liang, Junlin Li, Zhen Han, Xihao Wang, Zhongyuan Wang, Bin Mei
Abstract
Electrocautery or lasers will inevitably generate surgical smoke, which hinders the visual guidance of laparoscopic videos for surgical procedures. The surgical smoke can be classified into different types based on its motion patterns, leading to distinctive spatio-temporal characteristics across smoky laparoscopic videos. However, existing desmoking methods fail to account for such smoke-type-specific distinctions. Therefore, we propose the first Smoke-Type-Aware Laparoscopic Video Desmoking Network (STANet) by introducing two smoke types: Diffusion Smoke and Ambient Smoke. Specifically, a smoke mask segmentation sub-network is designed to jointly conduct smoke mask and smoke type predictions based on the attention-weighted mask aggregation, while a smokeless video reconstruction sub-network is proposed to perform specially desmoking on smoky features guided by two types of smoke mask. To address the entanglement challenges of two smoke types, we further embed a coarse-to-fine disentanglement module into the mask segmentation sub-network, which yields more accurate disentangled masks through the smoke-type-aware cross attention between non-entangled and entangled regions. In addition, we also construct the first large-scale synthetic video desmoking dataset with smoke type annotations. Extensive experiments demonstrate that our method not only outperforms state-of-the-art approaches in quality evaluations, but also exhibits superior generalization across multiple downstream surgical tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cb051d69-fdce-4bdc-8ef1-743f0c3e8665Builds on13
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- Image Dehazing Transformer with Transmission-Aware 3D Position EmbeddingChunle Guo, Qixin Yan, Saeed Anwar, Runmin Cong et al.CVPR 2022 · 464 citations
- SurgicalSAM: Efficient Class Promptable Surgical Instrument SegmentationWenxi Yue, Jing Zhang, Kun Hu, Yong Xia et al.AAAI 2024 · 142 citations
- Guided Real Image Dehazing Using YCbCr Color SpaceWenxuan Fang, Junkai Fan, Yu Zheng, Jiangwei Weng et al.AAAI 2025 · 46 citations
- Snow Removal in Video: A New Dataset and A Novel MethodHaoyu Chen, Jingjing Ren, Jinjin Gu, Hongtao Wu et al.ICCV 2023 · 40 citations
Related papers
- Benchmarking Endoscopic Surgical Image Restoration and BeyondJialun Pei, Diandian Guo, Donghui Yang, Zhixi Li et al.CVPR 2026
- Synergistic Bleeding Region and Point Detection in Laparoscopic Surgical VideosJialun Pei, Zhangjun Zhou, Diandian Guo, Zhixi Li et al.CVPR 2026 · 6 citations
- FoSp: Focus and Separation Network for Early Smoke SegmentationLujian Yao, Haitao Zhao, Jingchao Peng, Zhongze Wang et al.AAAI 2024 · 17 citations
- DAM-VSR: Disentanglement of Appearance and Motion for Video Super-ResolutionZhe Kong, Le Li, Yong Zhang, Feng Gao et al.SIGGRAPH 2025 · 6 citations
- Mitigating Surgical Data Imbalance with Dual-Prediction Video Diffusion ModelDanush Kumar Venkatesh, Adam Schmidt, Muhammad Abdullah Jamal, Omid MohareriICML 2026 · 1 citation
