How High is ‘High’? Rethinking the Roles of Dimensionality in Topological Data Analysis and Manifold Learning
Hannah Sansford, Nick Whiteley, Patrick Rubin-Delanchy
摘要
High-dimensionality of data is often regarded as a fundamental statistical impediment in Machine Learning and AI. The purpose of this paper is to clarify, on the contrary, when and how high-dimensionality may be beneficial. In the setting of a general random function model of data we delineate between three notions of dimensionality: effective dimension , measuring total variability across feature directions; correlation rank , measuring functional complexity across samples; and latent intrinsic dimension of manifold structure hidden in data. Via a generalized Hanson-Wright inequality, we show that increasing drives a blessing of dimensionality phenomenon, whereby data dot-products concentrate about their expectations. In turn, we show that, under mild continuity assumptions (ensuring that features bring additional information as dimension grows), persistence diagrams recover latent homology when as . Informed by our theory, we revisit the ground-breaking neuroscience discovery of toroidal structure in grid-cell activity made by Gardner et al. (2022): our findings provide the first empirical evidence that this structure is isometric to a flat torus model of physical space, suggesting that grid cell activity conveys a geometrically faithful representation of the real world.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum 等ICLR 2021 · 被引用 381 次
- The Computational Advantage of Depth in Learning High-Dimensional Hierarchical TargetsYatin Dandi, Luca Pesce, Lenka Zdeborová, Florent KrzakalaNeurIPS 2025 · 被引用 3 次
- Learning Multi-Index Models with Neural Networks via Mean-Field Langevin DynamicsAlireza Mousavi-Hosseini, Denny Wu, Murat A. ErdogduICLR 2025
- Local Intrinsic Dimension of Representations Predicts Alignment and Generalization in AI Models and Human BrainJunjie Yu, Wenxiao Ma, Chen Wei, Jianyu Zhang 等ICML 2026 · 被引用 2 次
- Intrinsic Dimension, Persistent Homology and Generalization in Neural NetworksTolga Birdal, Aaron Lou, Leonidas J. Guibas, Umut SimsekliNeurIPS 2021 · 被引用 94 次
