Harmony: A Generic Unsupervised Approach for Disentangling Semantic Content from Parameterized Transformations
Mostofa Rafid Uddin, Gregory Howe, Xiangrui Zeng, Min Xu
摘要
In many real-life image analysis applications, particularly in biomedical research domains, the objects of interest undergo multiple transformations that alters their visual properties while keeping the semantic content unchanged. Disentangling images into semantic content factors and transformations can provide significant benefits into many domain-specific image analysis tasks. To this end, we propose a generic unsupervised framework, Harmony, that simultaneously and explicitly disentangles semantic content from multiple parameterized transformations. Harmony leverages a simple cross-contrastive learning framework with multiple explicitly parameterized latent representations to disentangle content from transformations. To demonstrate the efficacy of Harmony, we apply it to disentangle image semantic content from several parameterized transformations (rotation, translation, scaling, and contrast). Harmony achieves significantly improved disentanglement over the baseline models on several image datasets of diverse domains. With such disentanglement, Harmony is demonstrated to incentivize bioimage analysis research by modeling structural heterogeneity of macromolecules from cryo-ET images and learning transformation-invariant representations of protein particles from single-particle cryo-EM images. Harmony also performs very well in disentangling content from 3D transformations and can perform coarse and fast alignment of 3D cryo-ET subtomograms. Therefore, Harmony is generalizable to many other imaging domains and can potentially be extended to domains beyond imaging as well.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Amortized Inference for Heterogeneous Reconstruction in Cryo-EMAxel Levy, Gordon Wetzstein, Julien N. P. Martel, Frédéric Poitevin 等NeurIPS 2022 · 被引用 57 次
- Unsupervised Identification of Protein Compositions and Conformations Via Implicit Content-Transformation DisentanglementMostofa Rafid Uddin, Jana Armouti, Min XuICCV 2025
它引用的顶会 Paper4
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Gum-Net: Unsupervised Geometric Matching for Fast and Accurate 3D Subtomogram Image Alignment and AveragingXiangrui Zeng, Min XuCVPR 2020
相关 Paper
- Rotation and Translation Invariant Representation Learning with Implicit Neural RepresentationsSehyun Kwon, Joo Young Choi, Ernest K. RyuICML 2023 · 被引用 5 次
- Unsupervised Object Representation Learning using Translation and Rotation Group Equivariant VAEAlireza Nasiri, Tristan BeplerNeurIPS 2022 · 被引用 18 次
- Image Harmonization with TransformerZonghui Guo, Dongsheng Guo, Haiyong Zheng, Zhaorui Gu 等ICCV 2021 · 被引用 95 次
- End-to-end robust joint unsupervised image alignment and clusteringXiangrui Zeng, Gregory Howe, Min XuICCV 2021 · 被引用 12 次
- DRANet: Disentangling Representation and Adaptation Networks for Unsupervised Cross-Domain AdaptationSeunghun Lee, Sunghyun Cho, Sunghoon ImCVPR 2021
