Representing Part-Whole Hierarchies in Foundation Models by Learning Localizability, Composability, and Decomposability from Anatomy via Self-Supervision
Mohammad Reza Hosseinzadeh Taher, Michael B. Gotway, Jianming Liang
摘要
Humans effortlessly interpret images by parsing them into part-whole hierarchies; deep learning excels in learning multi-level feature spaces, but they often lack explicit coding of part-whole relations, a prominent property of medical imaging. To overcome this limitation, we introduce Adam-v2, a new self-supervised learning framework extending Adam [79] by explicitly incorporating part-whole hierarchies into its learning objectives through three key branches: (1) Localizability, acquiring discriminative representations to distinguish different anatomical patterns; (2) Composability, learning each anatomical structure in a parts-to-whole manner; and (3) Decomposability, comprehending each anatomical structure in a whole-to-parts manner. Experimental results across 10 tasks, compared to 11 baselines in zero-shot, few-shot transfer, and full fine-tuning settings, showcase Adam-v2's superior performance over large-scale medical models and existing SSL methods across diverse downstream tasks. The higher generality and robustness of Adam-v2's representations originate from its explicit construction of hierarchies for distinct anatomical structures from unlabeled medical images. Adam-v2 preserves a semantic balance of anatomical diversity and harmony in its embedding, yielding representations that are both generic and semantically meaningful, yet overlooked in existing SSL methods. All code and pretrained models are available at GitHub.com/JLiangLab/Eden.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- AFiRe: Anatomy-Driven Self-Supervised Learning for Fine-Grained Representation in Radiographic ImagesYihang Liu, Lianghua He, Ying Wen, Longzhen Yang 等AAAI 2025 · 被引用 7 次
- CoSMIC: Continual Self-Supervised Learning for Multi-Domain Medical Imaging Via Conditional Mutual Information MaximizationYihang Liu, Ying Wen, Longzhen Yang, Lianghua He 等ICCV 2025 · 被引用 2 次
- A Flag Decomposition for Hierarchical DatasetsNathan Mankovich, Ignacio Santamaría, Gustau Camps-Valls, Tolga BirdalCVPR 2025
- M-IDoL: Information Decomposition for Modality-Specific and Diverse Representation Learning in Medical Foundation ModelYihang Liu, Longzhen Yang, Jiaxiong Yang, Ying Wen 等ICML 2026
- CheXWorld: Exploring Image World Modeling for Radiograph Representation LearningYang Yue, Yulin Wang, Chenxin Tao, Pan Liu 等CVPR 2025
它引用的顶会 Paper54
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
相关 Paper
- Anatomical Invariance Modeling and Semantic Alignment for Self-supervised Learning in 3D Medical Image AnalysisYankai Jiang, Mingze Sun, Heng Guo, Xiaoyu Bai 等ICCV 2023 · 被引用 38 次
- SAMora: Enhancing SAM through Hierarchical Self-Supervised Pre-Training for Medical ImagesShuhang Chen, Hangjie Yuan, Pengwei Liu, Hanxue Gu 等ICCV 2025 · 被引用 1 次
- DiRA: Discriminative, Restorative, and Adversarial Learning for Self-supervised Medical Image AnalysisFatemeh Haghighi, Mohammad Reza Hosseinzadeh Taher, Michael B. Gotway, Jianming LiangCVPR 2022 · 被引用 85 次
- Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-SupervisionYunhe Gao, Yabin Zhang, Chong Wang, Jiaming Liu 等CVPR 2026
- Structure-Aware Semantic Discrepancy and Consistency for 3D Medical Image Self-Supervised LearningTan Pan, Zhaorui Tan, Kaiyu Guo, Dongli Xu 等ICCV 2025 · 被引用 2 次
