Do Topological Characteristics Help in Knowledge Distillation?
Jungeun Kim, Junwon You, Dongjin Lee, Ha Young Kim, Jae-Hun Jung
Abstract
Knowledge distillation (KD) aims to transfer knowledge from larger (teacher) to smaller (student) networks. Previous studies focus on pointto-point or pairwise relationships in embedding features as knowledge and struggle to efficiently transfer relationships of complex latent spaces. To tackle this issue, we propose a novel KD method called TopKD, which considers the global topology of the latent spaces. We define global topology knowledge using the persistence diagram (PD) that captures comprehensive geometric structures such as shape of distribution, multiscale structure and connectivity, and the topology distillation loss for teaching this knowledge. To make the PD transferable within reasonable computational time, we employ approximated persistence images of PDs. Through experiments, we support the benefits of using global topology as knowledge and demonstrate the potential of TopKD. Code is available at https://github.com/ jekim5418/TopKD
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 87cdd738-2afb-4758-ad84-14898747a8ecCited by top-tier papers5
- Persistence Homology Distillation for Semi-supervised Continual LearningYan Fan, Yu Wang, Pengfei Zhu, Dongyue Chen et al.NeurIPS 2024 · 12 citations
- Evidential Knowledge DistillationLiangyu Xiang, Junyu Gao, Changsheng XuICCV 2025 · 6 citations
- Neural Collapse Inspired Knowledge DistillationShuoxi Zhang, Zijian Song, Kun HeAAAI 2025 · 2 citations
- Heterogeneous Complementary DistillationLiuchi Xu, Hao Zheng, Lu Wang, Lisheng Xu et al.AAAI 2026
- Topology-aware Knowledge Preservation for Class-Incremental LearningHan Zang, Yongfeng Dong, Linhao Li, Liang Yang et al.AAAI 2026
Builds on22
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 1,305 citations
- Similarity-Preserving Knowledge DistillationFrederick Tung, Greg MoriICCV 2019 · 1,214 citations
- Decoupled Knowledge DistillationBorui Zhao, Quan Cui, Renjie Song, Yiyu Qiu et al.CVPR 2022 · 835 citations
- A Comprehensive Overhaul of Feature DistillationByeongho Heo, Jeesoo Kim, Sangdoo Yun, Hyojin Park et al.ICCV 2019 · 727 citations
- Correlation Congruence for Knowledge DistillationBaoyun Peng, Xiao Jin, Dongsheng Li, Shunfeng Zhou et al.ICCV 2019 · 625 citations
Related papers
- Distilling Knowledge From Graph Convolutional NetworksYiding Yang, Jiayan Qiu, Mingli Song, Dacheng Tao et al.CVPR 2020
- Topology Distillation for Recommender SystemSeongKu Kang, Junyoung Hwang, Wonbin Kweon, Hwanjo YuKDD 2021 · 34 citations
- Distilling Holistic Knowledge with Graph Neural NetworksSheng Zhou, Yucheng Wang, Defang Chen, Jiawei Chen et al.ICCV 2021 · 69 citations
- Cross-Image Relational Knowledge Distillation for Semantic SegmentationChuanguang Yang, Helong Zhou, Zhulin An, Xue Jiang et al.CVPR 2022 · 228 citations
- Learning topology-preserving data representationsIlya Trofimov, Daniil Cherniavskii, Eduard Tulchinskii, Nikita Balabin et al.ICLR 2023 · 2 citations
