Iteratively-Refined Interactive 3D Medical Image Segmentation With Multi-Agent Reinforcement Learning
Xuan Liao, Wenhao Li, Qisen Xu, Xiangfeng Wang, Bo Jin, Xiaoyun Zhang, Yanfeng Wang, Ya Zhang
Abstract
Existing automatic 3D image segmentation methods usually fail to meet the clinic use. Many studies have explored an interactive strategy to improve the image segmentation performance by iteratively incorporating user hints. However, the dynamic process for successive interactions is largely ignored. We here propose to model the dynamic process of iterative interactive image segmentation as a Markov decision process (MDP) and solve it with reinforcement learning (RL). Unfortunately, it is intractable to use single-agent RL for voxel-wise prediction due to the large exploration space. To reduce the exploration space to a tractable size, we treat each voxel as an agent with a shared voxel-level behavior strategy so that it can be solved with multi-agent reinforcement learning. An additional advantage of this multi-agent model is to capture the dependency among voxels for segmentation task. Meanwhile, to enrich the information of previous segmentations, we reserve the prediction uncertainty in the state space of MDP and derive an adjustment action space leading to a more precise and finer segmentation. In addition, to improve the efficiency of exploration, we design a relative cross-entropy gain-based reward to update the policy in a constrained direction. Experimental results on various medical datasets have shown that our method significantly outperforms existing state-ofthe-art methods, with the advantage of fewer interactions and a faster convergence.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6f0cf4d4-ca30-4e11-bb82-720214435d7dCited by top-tier papers12
- FocusCut: Diving into a Focus View in Interactive SegmentationZheng Lin, Zheng-Peng Duan, Zhao Zhang, Chun-Le Guo et al.CVPR 2022 · 61 citations
- Neural Volumetric Object SelectionZhongzheng Ren, Aseem Agarwala, Bryan C. Russell, Alexander G. Schwing et al.CVPR 2022 · 41 citations
- Lane2Seq: Towards Unified Lane Detection via Sequence GenerationKunyang ZhouCVPR 2024 · 28 citations
- Weakly Supervised 3D Semantic Segmentation Using Cross-Image Consensus and Inter-Voxel Affinity RelationsXiaoyu Zhu, Jeffrey Chen, Xiangrui Zeng, Junwei Liang et al.ICCV 2021 · 20 citations
- Dynamic Policy-Driven Adaptive Multi-Instance Learning for Whole Slide Image ClassificationTingting Zheng, Kui Jiang, Hongxun YaoCVPR 2024 · 16 citations
Related papers
- Learning To Recommend Frame for Interactive Video Object Segmentation in the WildZhaoyuan Yin, Jia Zheng, Weixin Luo, Shenhan Qian et al.CVPR 2021
- MARL-MambaContour: Unleashing Multi-Agent Deep Reinforcement Learning for Active Contour Optimization in Medical Image SegmentationRuicheng Zhang, Yu Sun, Zeyu Zhang, Jinai Li et al.ACM MM 2025 · 1 citation
- IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and SegmentationYankai Jiang, Qiaoru Li, Binlu Xu, Haoran Sun et al.CVPR 2026 · 9 citations
- PixelSeg: Pixel-by-Pixel Stochastic Semantic Segmentation for Ambiguous Medical ImagesWei Zhang, Xiaohong Zhang, Sheng Huang, Yuting Lu et al.ACM MM 2022 · 10 citations
- ColorRL: Reinforced Coloring for End-to-End Instance SegmentationTuan Tran Anh, Khoa Nguyen-Tuan, Tran Minh Quan, Won-Ki JeongCVPR 2021
