Persistent Topological Features in Large Language Models
Yuri Gardinazzi, Karthik Viswanathan, Giada Panerai, Alessio Ansuini, Alberto Cazzaniga, Matteo Biagetti
Abstract
Understanding the decision-making processes of large language models is critical given their widespread applications. To achieve this, we aim to connect a formal mathematical framework-zigzag persistence from topological data analysis -with practical and easily applicable algorithms. Zigzag persistence is particularly effective for characterizing data as it dynamically transforms across model layers. Within this framework, we introduce topological descriptors that measure how topological features, p-dimensional holes, persist and evolve throughout the layers. Unlike methods that assess each layer individually and then aggregate the results, our approach directly tracks the full evolutionary path of these features. This offers a statistical perspective on how prompts are rearranged and their relative positions changed in the representation space, providing insights into the system's operation as an integrated whole. To demonstrate the expressivity and applicability of our framework, we highlight how sensitive these descriptors are to different models and a variety of datasets. As a showcase application to a downstream task, we use zigzag persistence to establish a criterion for layer pruning, achieving results comparable to state-ofthe-art methods while preserving the system-level perspective.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f48ba28c-56ca-46d8-8259-0f8285e059a1Cited by top-tier papers4
- The Shape of Adversarial Influence: Characterizing LLM Latent Spaces with Persistent HomologyAideen Fay, Inés García-Redondo, Qiquan Wang, Haim Dubossarsky et al.ICLR 2026 · 8 citations
- Large Vision-Language Models Get Lost in AttentionGongli Xi, Ye Tian, Mengyu Yang, Huahui Yi et al.ICML 2026 · 4 citations
- When Annotators Disagree, Topology Explains: Mapper, a Topological Tool for Exploring Text Embedding Geometry and AmbiguityNisrine Rair, Alban Goupil, Valeriu Vrabie, Emmanuel ChochoyEMNLP 2025
- KDP: Simplifying Representation Dynamics in Kernel SpaceZeyu Ma, Wanying Wang, Guchu Zou, Mingang Chen et al.ICLR 2026
Builds on13
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou et al.ICLR 2021 · 7,905 citations
- Reducing Transformer Depth on Demand with Structured DropoutAngela Fan, Edouard Grave, Armand JoulinICLR 2020 · 695 citations
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- Intrinsic Dimension Estimation for Robust Detection of AI-Generated TextsEduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva, Daniil Cherniavskii et al.NeurIPS 2023 · 163 citations
- The geometry of hidden representations of large transformer modelsLucrezia Valeriani, Diego Doimo, Francesca Cuturello, Alessandro Laio et al.NeurIPS 2023 · 148 citations
Related papers
- Large Language Models as Topological Thinkers: A Benchmark on Graph Persistent HomologyHao Li, Hao Wan, Yixue Huang, Yuzhou Chen et al.ICML 2026
- Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitionsboyu shi, Chang Liu, Chuanbao Gao, Xu Yang et al.ICML 2026 · 1 citation
- Topological Zigzag Spaghetti for Diffusion-based Generation and Prediction on GraphsYuzhou Chen, Yulia R. GelICLR 2025
- Experimental Observations of the Topology of Convolutional Neural Network ActivationsEmilie Purvine, Davis Brown, Brett A. Jefferson, Cliff A. Joslyn et al.AAAI 2023 · 21 citations
- Text summarization via global structure awarenessJiaquan Zhang, Chaoning Zhang, Shuxu Chen, Yibei Liu et al.ICLR 2026 · 14 citations
