Pooling Image Datasets with Multiple Covariate Shift and Imbalance
Sotirios Panagiotis Chytas, Vishnu Suresh Lokhande, Vikas Singh
摘要
Small sample sizes are common in many disciplines, which necessitates pooling roughly similar datasets across multiple institutions to study weak but relevant associations between images and disease outcomes. Such data often manifest shift/imbalance in covariates (i.e., secondary non-imaging data). Controlling for such nuisance variables is common within standard statistical analysis, but the ideas do not directly apply to overparameterized models. Consequently, recent work has shown how strategies from invariant representation learning provides a meaningful starting point, but the current repertoire of methods is limited to accounting for shifts/imbalances in just a couple of covariates at a time. In this paper, we show how viewing this problem from the perspective of Category theory provides a simple and effective solution that completely avoids elaborate multi-stage training pipelines that would otherwise be needed. We show the effectiveness of this approach via extensive experiments on real datasets. Further, we discuss how this style of formulation offers a unified perspective on at least 5+ distinct problem settings, from self-supervised learning to matching problems in 3D reconstruction. The code is available at https://github.com/SPChytas/CatHarm .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 被引用 1,025 次
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black 等ICLR 2020 · 被引用 298 次
- Topological AutoencodersMichael Moor, Max Horn, Bastian Rieck, Karsten M. BorgwardtICML 2020 · 被引用 192 次
相关 Paper
- Equivariance Allows Handling Multiple Nuisance Variables When Analyzing Pooled Neuroimaging DatasetsVishnu Suresh Lokhande, Rudrasis Chakraborty, Sathya N. Ravi, Vikas SinghCVPR 2022 · 被引用 1 次
- Do deep networks transfer invariances across classes?Allan Zhou, Fahim Tajwar, Alexander Robey, Tom Knowles 等ICLR 2022 · 被引用 19 次
- Stabilizing In-Context Multi-Source Domain Adaptation for Biomedical Images Through ControlsAna Sanchez Fernandez, Thomas Pinetz, Werner Zellinger, Günter KlambauerICML 2026
- Invariant and Transportable Representations for Anti-Causal Domain ShiftsYibo Jiang, Victor VeitchNeurIPS 2022 · 被引用 50 次
- An Investigation of Representation and Allocation Harms in Contrastive LearningSubha Maity, Mayank Agarwal, Mikhail Yurochkin, Yuekai SunICLR 2024 · 被引用 2 次
