Temporal Context Matters: Enhancing Single Image Prediction with Disease Progression Representations
Aishik Konwer, Xuan Xu, Joseph Bae, Chao Chen, Prateek Prasanna
摘要
Clinical outcome or severity prediction from medical images has largely focused on learning representations from single-timepoint or snapshot scans. It has been shown that disease progression can be better characterized by temporal imaging. We therefore hypothesized that outcome predictions can be improved by utilizing the disease progression informationfrom sequential images. We present a deep learning approach that leverages temporal progression information to improve clinical outcome predictions from single-timepoint images. In our method, a self-attention based Temporal Convolutional Network (TCN) is used to learn a representation that is most reflective of the disease trajectory. Meanwhile, a Vision Transformer is pretrained in a self-supervised fashion to extract features from single-timepoint images. The key contribution is to design a recalibration module that employs maximum mean discrepancy loss (MMD) to align distributions of the above two contextual representations. We train our system to predict clinical outcomes and severity grades from single-timepoint images. Experiments on chest and osteoarthritis radiography datasets demonstrate that our approach outperforms other state-of-the-art techniques.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Enhancing Modality-Agnostic Representations via Meta-learning for Brain Tumor SegmentationAishik Konwer, Xiaoling Hu, Joseph Bae, Xuan Xu 等ICCV 2023 · 被引用 23 次
- Enhancing SAM with Efficient Prompting and Preference Optimization for Semi-supervised Medical Image SegmentationAishik Konwer, Zhijian Yang, Erhan Bas, Cao Xiao 等CVPR 2025
它引用的顶会 Paper11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Big Self-Supervised Models Advance Medical Image ClassificationShekoofeh Azizi, Basil Mustafa, Fiona Ryan, Zachary Beaver 等ICCV 2021 · 被引用 695 次
- Multimodal Co-Attention Transformer for Survival Prediction in Gigapixel Whole Slide ImagesRichard J. Chen, Ming Y. Lu, Wei-Hung Weng, Tiffany Y. Chen 等ICCV 2021 · 被引用 369 次
- Probabilistic Regression for Visual TrackingMartin Danelljan, Luc Van Gool, Radu TimofteCVPR 2020
相关 Paper
- Learning to Exploit Temporal Structure for Biomedical Vision-Language ProcessingShruthi Bannur, Stephanie L. Hyland, Qianchu Liu, Fernando Pérez-García 等CVPR 2023
- Temporal Inversion for Learning Interval Change in Chest X-RaysHanbin Ko, Kyeongmin Jeon, Doowoong Choi, Chang Min ParkCVPR 2026 · 被引用 3 次
- Unlocking the Power of Spatial and Temporal Information in Medical Multimodal Pre-trainingJinxia Yang, Bing Su, Xin Zhao, Ji-Rong WenICML 2024 · 被引用 13 次
- Medical Vision-Language Pretraining with LLM-Guided Temporal SupervisionLiang Bai, Zhi Wang, Huimin Yan, Xian YangAAAI 2026
- Scaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT TransformersCris Claessens, Christiaan Viviers, Giacomo D'Amicantonio, Egor Bondarev 等CVPR 2026 · 被引用 6 次
