Persistent Topological Features in Large Language Models
Yuri Gardinazzi, Karthik Viswanathan, Giada Panerai, Alessio Ansuini, Alberto Cazzaniga, Matteo Biagetti
摘要
Understanding the decision-making processes of large language models is critical given their widespread applications. To achieve this, we aim to connect a formal mathematical framework-zigzag persistence from topological data analysis -with practical and easily applicable algorithms. Zigzag persistence is particularly effective for characterizing data as it dynamically transforms across model layers. Within this framework, we introduce topological descriptors that measure how topological features, p-dimensional holes, persist and evolve throughout the layers. Unlike methods that assess each layer individually and then aggregate the results, our approach directly tracks the full evolutionary path of these features. This offers a statistical perspective on how prompts are rearranged and their relative positions changed in the representation space, providing insights into the system's operation as an integrated whole. To demonstrate the expressivity and applicability of our framework, we highlight how sensitive these descriptors are to different models and a variety of datasets. As a showcase application to a downstream task, we use zigzag persistence to establish a criterion for layer pruning, achieving results comparable to state-ofthe-art methods while preserving the system-level perspective.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- The Shape of Adversarial Influence: Characterizing LLM Latent Spaces with Persistent HomologyAideen Fay, Inés García-Redondo, Qiquan Wang, Haim Dubossarsky 等ICLR 2026 · 被引用 8 次
- Large Vision-Language Models Get Lost in AttentionGongli Xi, Ye Tian, Mengyu Yang, Huahui Yi 等ICML 2026 · 被引用 4 次
- When Annotators Disagree, Topology Explains: Mapper, a Topological Tool for Exploring Text Embedding Geometry and AmbiguityNisrine Rair, Alban Goupil, Valeriu Vrabie, Emmanuel ChochoyEMNLP 2025
- KDP: Simplifying Representation Dynamics in Kernel SpaceZeyu Ma, Wanying Wang, Guchu Zou, Mingang Chen 等ICLR 2026
它引用的顶会 Paper13
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou 等ICLR 2021 · 被引用 7,905 次
- Reducing Transformer Depth on Demand with Structured DropoutAngela Fan, Edouard Grave, Armand JoulinICLR 2020 · 被引用 695 次
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum 等ICLR 2021 · 被引用 381 次
- Intrinsic Dimension Estimation for Robust Detection of AI-Generated TextsEduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva, Daniil Cherniavskii 等NeurIPS 2023 · 被引用 163 次
- The geometry of hidden representations of large transformer modelsLucrezia Valeriani, Diego Doimo, Francesca Cuturello, Alessandro Laio 等NeurIPS 2023 · 被引用 148 次
相关 Paper
- Large Language Models as Topological Thinkers: A Benchmark on Graph Persistent HomologyHao Li, Hao Wan, Yixue Huang, Yuzhou Chen 等ICML 2026
- Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitionsboyu shi, Chang Liu, Chuanbao Gao, Xu Yang 等ICML 2026 · 被引用 1 次
- Topological Zigzag Spaghetti for Diffusion-based Generation and Prediction on GraphsYuzhou Chen, Yulia R. GelICLR 2025
- Experimental Observations of the Topology of Convolutional Neural Network ActivationsEmilie Purvine, Davis Brown, Brett A. Jefferson, Cliff A. Joslyn 等AAAI 2023 · 被引用 21 次
- Text summarization via global structure awarenessJiaquan Zhang, Chaoning Zhang, Shuxu Chen, Yibei Liu 等ICLR 2026 · 被引用 14 次
