Local Spatiotemporal Representation Learning for Longitudinally-consistent Neuroimage Analysis
Mengwei Ren, Neel Dey, Martin Styner, Kelly N. Botteron, Guido Gerig
Abstract
Recent self-supervised advances in medical computer vision exploit the global and local anatomical self-similarity for pretraining prior to downstream tasks such as segmentation. However, current methods assume i.i.d. image acquisition, which is invalid in clinical study designs where follow-up longitudinal scans track subject-specific temporal changes. Further, existing self-supervised methods for medically-relevant image-to-image architectures exploit only spatial or temporal self-similarity and do so via a loss applied only at a single image-scale, with naive multi-scale spatiotemporal extensions collapsing to degenerate solutions. To these ends, this paper makes two contributions: (1) It presents a local and multi-scale spatiotemporal representation learning method for image-to-image architectures trained on longitudinal images. It exploits the spatiotemporal self-similarity of learned multi-scale intra-subject image features for pretraining and develops several feature-wise regularizations that avoid degenerate representations; (2) During finetuning, it proposes a surprisingly simple self-supervised segmentation consistency regularization to exploit intra-subject correlation. Benchmarked across various segmentation tasks, the proposed framework outperforms both well-tuned randomly-initialized baselines and current self-supervised techniques designed for both i.i.d. and longitudinal datasets. These improvements are demonstrated across both longitudinal neurodegenerative adult MRI and developing infant brain MRI and yield both higher performance and longitudinal consistency. 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Sequential Multi-Dimensional Self-Supervised Learning for Clinical Time SeriesAniruddh Raghu, Payal Chandak, Ridwan Alam, John V. Guttag et al.ICML 2023 · 18 citations
- Keypoint-Augmented Self-Supervised Learning for Medical Image Segmentation with Limited AnnotationZhangsihao Yang, Mengwei Ren, Kaize Ding, Guido Gerig et al.NeurIPS 2023 · 12 citations
- Learning Patient-Specific Disease Dynamics With Latent Flow Matching For Longitudinal Imaging GenerationHao Chen, Rui Yin, Yifan Chen, Qi Chen et al.ICLR 2026 · 11 citations
- Multimodal Disease Progression Modeling via Spatiotemporal Disentanglement and Multiscale AlignmentChen Liu, Wenfang Yao, Kejing Yin, William K. Cheung et al.NeurIPS 2025 · 4 citations
- Dual Meta-Learning with Longitudinally Generalized Regularization for One-Shot Brain Tissue Segmentation Across the Human LifespanYongheng Sun, Fan Wang, Jun Shu, Haifeng Wang et al.ICCV 2023 · 2 citations
Builds on19
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
Related papers
- Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical ImagingTan Pan, Shuhao Mei, Yixuan Sun, Kaiyu Guo et al.ICML 2026
- Autoregressive Sequence Modeling for 3D Medical Image RepresentationSiwen Wang, Churan Wang, Fei Gao, Lixian Su et al.AAAI 2025 · 5 citations
- Multi-modal Vision Pre-training for Medical Image AnalysisShaohao Rui, Lingzhi Chen, Zhenyu Tang, Lilong Wang et al.CVPR 2025
- An OpenMind for 3D Medical Vision Self-supervised LearningTassilo Wald, Constantin Ulrich, Jonathan Suprijadi, Sebastian Ziegler et al.ICCV 2025 · 5 citations
- Continual Self-Supervised Learning: Towards Universal Multi-Modal Medical Data Representation LearningYiwen Ye, Yutong Xie, Jianpeng Zhang, Ziyang Chen et al.CVPR 2024 · 30 citations
