Unsupervised Multi-Scale Segmentation of 3D Subcellular World with Stable Diffusion Foundation Model
Mostofa Rafid Uddin, H. M. Shadman Tabib, Thanh-Huy Nguyen, Kashish Gandhi, Min Xu
Abstract
We introduce an unsupervised approach for segmenting multiscale subcellular objects in 3D volumetric cryoelectron tomography (cryo-ET) images. To this end, we address key challenges such as lack of annotated data, large data volumes, high heterogeneity of subcellular shapes and sizes, and high inter-domain variability of cellular cryo-ET images across different experiments and contexts. Our method requires users to only select a small number of slabs from a few representative tomograms in the dataset. The core of our method is extracting features for the corresponding slabs, leveraging a Stable Diffusion foundation model pretrained on mostly natural images. The feature extraction is followed by a novel heuristic-based feature aggregation strategy, and adaptive thresholding to segment the aggregated features. The resulting masks are refined with pretrained CellPose to split composite regions, and then utilized as pseudo-ground truth for training supervised deep learning models. We validated our unsupervised foundation-model based pipeline on publicly available cryo-ET benchmark datasets, demonstrating performance that closely approximates expert human annotations. This fully automated, data-driven framework enables the mining of multi-scale subcellular patterns, paving the way for accelerated biological discoveries from large-scale cellular cryo-ET datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6ab3a520-552f-44e1-bf49-c2432728d95aBuilds on4
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Deciphering 'What' and 'Where' Visual Pathways from Spectral Clustering of Layer-Distributed Neural RepresentationsXiao Zhang, David Yunis, Michael MaireCVPR 2024 · 1 citation
- Cut and Learn for Unsupervised Object Detection and Instance SegmentationXudong Wang, Rohit Girdhar, Stella X. Yu, Ishan MisraCVPR 2023
Related papers
- End-to-end robust joint unsupervised image alignment and clusteringXiangrui Zeng, Gregory Howe, Min XuICCV 2021 · 12 citations
- Gum-Net: Unsupervised Geometric Matching for Fast and Accurate 3D Subtomogram Image Alignment and AveragingXiangrui Zeng, Min XuCVPR 2020
- Unsupervised Learning of Object-Centric Embeddings for Cell Instance Segmentation in Microscopy ImagesSteffen Wolf, Manan Lalit, Katie McDole, Jan FunkeICCV 2023 · 11 citations
- Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-SupervisionYunhe Gao, Yabin Zhang, Chong Wang, Jiaming Liu et al.CVPR 2026
- CryoGEN: Generative Energy-based Models for Cryogenic Electron Tomography ReconstructionYunfei Teng, Yuxuan Ren, Kai Chen, Xi Chen et al.ICLR 2025
