Building Robust Vision Encoders for Cross-Dataset Evaluation in Immunofluorescent Microscopy
Umar Marikkar, Syed Sameed Husain, Muhammad Awais, Sara Atito
摘要
Immunofluorescence (IF) images reveal detailed information about structures and functions at the subcellular level. However, unlike RGB images, IF datasets pose challenges for deep learning models due to their inconsistencies in channel count and configuration, stemming from varying staining protocols across laboratories and studies. Although existing approaches build channel-adaptive models for training, they do not perform evaluations across IF datasets with unseen channel configurations. To address this, we first introduce a biologically informed view of cellular image channels by grouping them into either context or concept, where we treat the context channels as a reference for the concept channels in the image. We leverage this view to propose Channel Conditioned Cell Representations (C3R), a framework that learns representations that transfers well to both in-distribution (ID) and out-of-distribution (OOD) datasets which contain same and different channel configurations, respectively. C3R is a two-fold framework comprising a channel-adaptive encoder architecture and a masked knowledge distillation training strategy, both built around the context-concept principle. We find that C3R outperforms existing benchmarks on both ID and OOD tasks, while yielding state-of-the-art results on frozen encoder evaluation on the CHAMMI benchmark. Our method opens a new pathway for cross-dataset generalization between IF datasets, with no need for retraining on unseen channel configurations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- ClimaX: A foundation model for weather and climateTung Nguyen, Johannes Brandstetter, Ashish Kapoor, Jayesh K. Gupta 等ICML 2023 · 被引用 426 次
- Image BERT Pre-training with Online TokenizerJinghao Zhou, Chen Wei, Huiyu Wang, Wei Shen 等ICLR 2022 · 被引用 287 次
相关 Paper
- CHAMMI-75: Pre-training multi-channel models with heterogeneous microscopy imagesVidit Agrawal, John Peters, Tyler N. Thompson, Mohammad V. Sanian 等ICLR 2026 · 被引用 3 次
- ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy ImageNicolas Bourriez, Ihab Bendidi, Ethan Cohen, Gabriel Watkinson 等CVPR 2024
- Enhancing Feature Diversity Boosts Channel-Adaptive Vision TransformersChau Pham, Bryan A. PlummerNeurIPS 2024 · 被引用 15 次
- RepMode: Learning to Re-Parameterize Diverse Experts for Subcellular Structure PredictionDonghao Zhou, Chunbin Gu, Junde Xu, Furui Liu 等CVPR 2023
- Generalized and Invariant Single-Neuron In-Vivo Activity Representation LearningWei Wu, Yuxing Lu, Zhengrui Guo, Chi Zhang 等NeurIPS 2025
