A Probabilistic Graph Coupling View of Dimension Reduction
Hugues Van Assel, Thibault Espinasse, Julien Chiquet, Franck Picard
Abstract
Most popular dimension reduction (DR) methods like t-SNE and UMAP are based on minimizing a cost between input and latent pairwise similarities. Though widely used, these approaches lack clear probabilistic foundations to enable a full understanding of their properties and limitations. To that extent, we introduce a unifying statistical framework based on the coupling of hidden graphs using cross-entropy. These graphs induce a Markov random field dependency structure among the observations in both input and latent spaces. We show that existing pairwise similarity DR methods can be retrieved from our framework with particular choices of priors for the graphs. Moreover, this reveals that these methods relying on shift-invariant kernels suffer from a statistical degeneracy that explains poor performances in conserving coarse-grain dependencies. New links are drawn with PCA which appears as a non-degenerate graph coupling model. 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Contrastive Learning is Spectral Clustering on Similarity GraphZhiquan Tan, Yifan Zhang, Jingqin Yang, Yang YuanICLR 2024 · 34 citations
- SNEkhorn: Dimension Reduction with Symmetric Entropic AffinitiesHugues Van Assel, Titouan Vayer, Rémi Flamary, Nicolas CourtyNeurIPS 2023 · 14 citations
- Dimension Reduction with Locally Adjusted GraphsYingfan Wang, Yiyang Sun, Haiyang Huang, Cynthia RudinAAAI 2025 · 10 citations
- Lambda: Learning Matchable Prior For Entity Alignment with Unlabeled Dangling CasesHang Yin, Liyao Xiang, Dong Ding, Yuheng He et al.NeurIPS 2024 · 7 citations
- Fair Neighbor EmbeddingJaakko Peltonen, Wen Xu, Timo Nummenmaa, Jyrki NummenmaaICML 2023 · 7 citations
Related papers
- SpaceMAP: Visualizing High-Dimensional Data by Space ExpansionXinrui Zu, Qian TaoICML 2022 · 12 citations
- Contrastive Learning Can Find An Optimal Basis For Approximately View-Invariant FunctionsDaniel D. Johnson, Ayoub El Hanchi, Chris J. MaddisonICLR 2023 · 1 citation
- Hierarchical Nearest Neighbor Graph Embedding for Efficient Dimensionality ReductionM. Saquib Sarfraz, Marios Koulakis, Constantin Seibold, Rainer StiefelhagenCVPR 2022 · 12 citations
- Learning to Rank: How GNNs Solve Max-Clique and Sparse PCAElad Shoham, Omri Haber, Havana Rika, Dan VilenchikAAAI 2026
- Joint t-SNE for Comparable Projections of Multiple High-Dimensional DatasetsYinqiao Wang, Lu Chen, Jaemin Jo, Yunhai WangIEEE VIS 2021 · 33 citations
