Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation
Yunhe Gao
摘要
A major focus of clinical imaging workflow is disease diagnosis and management, leading to medical imaging datasets strongly tied to specific clinical objectives. This scenario has led to the prevailing practice of developing task-specific segmentation models, without gaining insights from widespread imaging cohorts. Inspired by the training program of medical radiology residents, we propose a shift towards universal medical image segmentation, a paradigm aiming to build medical image understanding foundation models by leveraging the diversity and commonality across clinical targets, body regions, and imaging modalities. Towards this goal, we develop Hermes, a novel context-prior learning approach to address the challenges of data heterogeneity and annotation differences in medical image segmentation. In a large collection of eleven diverse datasets (2,438 3D images) across five modalities (CT, PET, T1, T2 and cine MRI) and multiple body regions, we demonstrate the merit of the universal paradigm over the traditional paradigm on addressing multiple tasks within a single model. By exploiting the synergy across tasks, Hermes achieves state-of-theart performance on all testing datasets and shows superior model scalability. Results on two additional datasets reveals Hermes' strong performance for transfer learning, incremental learning, and generalization to downstream tasks. Hermes's learned priors demonstrate an appealing trait to reflect the intricate relations among tasks and modalities, which aligns with the established anatomical and imaging principles in radiology. The code is available 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- SimMLM: A Simple Framework for Multi-Modal Learning with Missing ModalitySijie Li, Chen Chen, Jungong HanICCV 2025 · 被引用 14 次
- Mamba Goes HoME: Hierarchical Soft Mixture-of-Experts for 3D Medical Image SegmentationSzymon Plotka, Gizem Mert, Maciej Chrabaszcz, Ewa Szczurek 等NeurIPS 2025 · 被引用 5 次
- K-Prism: A Knowledge-Guided and Prompt Integrated Universal Medical Image Segmentation ModelBangwei Guo, Yunhe Gao, Meng Ye, Difei Gu 等ICLR 2026 · 被引用 2 次
- Show and Segment: Universal Medical Image Segmentation via In-Context LearningYunhe Gao, Di Liu, Zhuowei Li, Yunsheng Li 等CVPR 2025
- Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-SupervisionYunhe Gao, Yabin Zhang, Chong Wang, Jiaming Liu 等CVPR 2026
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- Large Batch Optimization for Deep Learning: Training BERT in 76 minutesYang You, Jing Li, Sashank J. Reddi, Jonathan Hseu 等ICLR 2020 · 被引用 1,170 次
相关 Paper
- UniverSeg: Universal Medical Image SegmentationVictor Ion Butoi, Jose Javier Gonzalez Ortiz, Tianyu Ma, Mert R. Sabuncu 等ICCV 2023 · 被引用 163 次
- Medverse: A Universal Model for Full-Resolution 3D Medical Image Segmentation, Transformation and EnhancementJiesi Hu, Jianfeng Cao, Yanwu Yang, Chenfei Ye 等AAAI 2026 · 被引用 2 次
- Neuroverse3D: Developing in-Context Learning Universal Model for Neuroimaging in 3DJiesi Hu, Hanyang Peng, Yanwu Yang, Xutao Guo 等ICCV 2025
- One-Prompt to Segment All Medical ImagesJunde Wu, Min XuCVPR 2024 · 被引用 32 次
- SegAnyPET: Universal Promptable Segmentation from Positron Emission Tomography ImagesYichi Zhang, Le Xue, Wenbo Zhang, Lanlan Li 等ICCV 2025 · 被引用 7 次
