Estimating Dimensionality of Neural Representations from Finite Samples
Chanwoo Chun, Abdulkadir Canatar, SueYeon Chung, Daniel D. Lee
Abstract
The global dimensionality of a neural representation manifold provides rich insight into the computational process underlying both artificial and biological neural networks. However, all existing measures of global dimensionality are sensitive to the number of samples, i.e., the number of rows and columns of the sample matrix. We show that, in particular, the participation ratio of eigenvalues, a popular measure of global dimensionality, is highly biased with small sample sizes, and propose a bias-corrected estimator that is more accurate with finite samples and with noise. On synthetic data examples, we demonstrate that our estimator can recover the true known dimensionality. We apply our estimator to neural brain recordings, including calcium imaging, electrophysiological recordings, and fMRI data, and to the neural activations in a large language model, and show that our estimator is invariant to the sample size. Finally, our estimators can additionally be used to measure the local dimensionalities of curved neural manifolds by weighting the finite samples appropriately.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e9a5e8d8-a1fb-4626-8082-b8511c9affd7Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Learning Efficient Coding of Natural Images with Maximum Manifold Capacity RepresentationsThomas E. Yerxa, Yilun Kuang, Eero P. Simoncelli, SueYeon ChungNeurIPS 2023 · 44 citations
- A theory of high dimensional regression with arbitrary correlations between input features and target functions: sample complexity, multiple descent curves and a hierarchy of phase transitionsGabriel Mel, Surya GanguliICML 2021 · 24 citations
- Are Sparse Autoencoders Useful? A Case Study in Sparse ProbingSubhash Kantamneni, Joshua Engels, Senthooran Rajamanoharan, Max Tegmark et al.ICML 2025
- Layer by Layer: Uncovering Hidden Representations in Language ModelsOscar Skean, Md Rifat Arefin, Dan Zhao, Niket Patel et al.ICML 2025
Related papers
- Spectral Analysis of Representational Similarity with Limited NeuronsHyunmo Kang, Abdulkadir Canatar, SueYeon ChungNeurIPS 2025 · 4 citations
- Credit Assignment via Neural Manifold Noise CorrelationByungwoo Kang, Maceo Richards, Bernardo SabatiniICML 2026 · 1 citation
- Effective Minkowski Dimension of Deep Nonparametric Regression: Function Approximation and Statistical TheoriesZixuan Zhang, Minshuo Chen, Mengdi Wang, Wenjing Liao et al.ICML 2023 · 4 citations
- Local Intrinsic Dimension of Representations Predicts Alignment and Generalization in AI Models and Human BrainJunjie Yu, Wenxiao Ma, Chen Wei, Jianyu Zhang et al.ICML 2026 · 2 citations
- Dimensionality Mismatch Between Brains and Artificial Neural NetworksSantiago Galella, Maren H. Wehrheim, Matthias KaschubeNeurIPS 2025
