Reliable Measures of Spread in High Dimensional Latent Spaces
Anna C. Marbut, Katy McKinney-Bock, Travis J. Wheeler
摘要
Understanding geometric properties of the latent spaces of natural language processing models allows the manipulation of these properties for improved performance on downstream tasks. One such property is the amount of data spread in a model's latent space, or how fully the available latent space is being used. We demonstrate that the commonly used measures of data spread, average cosine similarity and a partition function min/max ratio I(V), do not provide reliable metrics to compare the use of latent space across data distributions. We propose and examine six alternative measures of data spread, all of which improve over these current metrics when applied to seven synthetic data distributions. Of our proposed measures, we recommend one principal component-based measure and one entropy-based measure that provide reliable, relative measures of spread and can be used to compare models of different sizes and dimensionalities. One such geometric property is a quantification of how
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- NerVE: Nonlinear Eigenspectrum Dynamics in LLM Feed-Forward NetworksNandan Kumar Jha, Brandon ReagenICLR 2026 · 被引用 4 次
- Spectral Scaling Laws in Language Models: emphHow Effectively Do Feed-Forward Networks Use Their Latent Space?Nandan Kumar Jha, Brandon ReagenEMNLP 2025
它引用的顶会 Paper4
- Attention is Not Only a Weight: Analyzing Transformers with Vector NormsGoro Kobayashi, Tatsuki Kuribayashi, Sho Yokoi, Kentaro InuiEMNLP 2020 · 被引用 138 次
- Improving Neural Language Generation with Spectrum ControlLingxiao Wang, Jing Huang, Kevin Huang, Ziniu Hu 等ICLR 2020 · 被引用 94 次
- IsoBN: Fine-Tuning BERT with Isotropic Batch NormalizationWenxuan Zhou, Bill Yuchen Lin, Xiang RenAAAI 2021 · 被引用 29 次
- Embedding Compression with Isotropic Iterative QuantizationSiyu Liao, Jie Chen, Yanzhi Wang, Qinru Qiu 等AAAI 2020 · 被引用 16 次
相关 Paper
- Metric Space Magnitude for Evaluating the Diversity of Latent RepresentationsKatharina Limbeck, Rayna Andreeva, Rik Sarkar, Bastian RieckNeurIPS 2024 · 被引用 27 次
- On the Predictive Power of Representation Dispersion in Language ModelsYanhong Li, Ming Li, Karen Livescu, Jiawei ZhouICLR 2026 · 被引用 7 次
- All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational QualityWilliam Timkey, Marten van SchijndelEMNLP 2021 · 被引用 59 次
- Mapping 1, 000+ Language Models via the Log-Likelihood VectorMomose Oyama, Hiroaki Yamagiwa, Yusuke Takase, Hidetoshi ShimodairaACL 2025
- Map of Encoders - Mapping Sentence Encoders using Quantum Relative EntropyGaifan Zhang, Danushka BollegalaACL 2026
