Geometric Constraints for Small Language Models to Understand and Expand Scientific Taxonomies
Liri Fang, Dongqi Fu, Jiawei Han, Jingrui He, Vetle I Torvik
摘要
Recent findings reveal that token embeddings of Large Language Models (LLMs) exhibit strong hyperbolicity. This insight motivates leveraging LLMs for scientific taxonomy tasks, where maintaining and expanding hierarchical knowledge structures is critical. Although potential, generally-trained LLMs face challenges in directly handling domain-specific taxonomies, including computational cost and hallucination. Meanwhile, Small Language Models (SLMs) provide a more economical alternative if empowered with proper knowledge transfer. In this work, we introduce SS-Mono (Structure-Semantic Monotonization), a novel pipeline that combines local taxonomy augmentation from LLMs, self-supervised fine-tuning of SLMs with geometric constraints, and LLM calibration. Our approach enables efficient and accurate taxonomy expansion across root, leaf, and intermediate nodes. Extensive experiments on both leaf and non-leaf expansion benchmarks demonstrate that a fine-tuned SLM (e.g., DistilBERT-base-110M) consistently outperforms frozen LLMs (e.g., GPT-4o, Gemma-2-9B) and domain-specific baselines. These findings highlight the promise of lightweight yet effective models for structured knowledge enrichment in scientific domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Hyperbolic Neural Networks++Ryohei Shimizu, Yusuke Mukuta, Tatsuya HaradaICLR 2021 · 被引用 791 次
- Harnessing Explanations: LLM-to-LM Interpreter for Enhanced Text-Attributed Graph Representation LearningXiaoxin He, Xavier Bresson, Thomas Laurent, Adam Perold 等ICLR 2024 · 被引用 151 次
- From Trees to Continuous Embeddings and Back: Hyperbolic Hierarchical ClusteringInes Chami, Albert Gu, Vaggos Chatziafratis, Christopher RéNeurIPS 2020 · 被引用 125 次
- TaxoExpan: Self-supervised Taxonomy Expansion with Position-Enhanced Graph Neural NetworkJiaming Shen, Zhihong Shen, Chenyan Xiong, Chi Wang 等WWW 2020 · 被引用 85 次
相关 Paper
- BLEND: Balanced and Leaf-Enhanced Dual Fine-Tuning for Taxonomy CompletionPankaj, Dhruv Kumar, Vinayak Abrol, Vikram GoyalWWW 2026
- TaxoLLaMA: WordNet-based Model for Solving Multiple Lexical Semantic TasksViktor Moskvoretskii, Ekaterina Neminova, Alina Lobanova, Alexander Panchenko 等ACL 2024 · 被引用 11 次
- Compress and Mix: Advancing Efficient Taxonomy Completion with Large Language ModelsHongyuan Xu, Yuhang Niu, Yanlong Wen, Xiaojie YuanWWW 2025 · 被引用 6 次
- OntoTune: Ontology-Driven Self-training for Aligning Large Language ModelsZhiqiang Liu, Chengtao Gan, Junjie Wang, Yichi Zhang 等WWW 2025 · 被引用 11 次
- TokAlign: Efficient Vocabulary Adaptation via Token AlignmentChong Li, Jiajun Zhang, Chengqing ZongACL 2025 · 被引用 7 次
