Distilling Relation Embeddings from Pretrained Language Models
Asahi Ushio, José Camacho-Collados, Steven Schockaert
Abstract
Pre-trained language models have been found to capture a surprisingly rich amount of lexical knowledge, ranging from commonsense properties of everyday concepts to detailed factual knowledge about named entities. Among others, this makes it possible to distill high-quality word vectors from pre-trained language models. However, it is currently unclear to what extent it is possible to distill relation embeddings, i.e. vectors that characterize the relationship between two words. Such relation embeddings are appealing because they can, in principle, encode relational knowledge in a more finegrained way than is possible with knowledge graphs. To obtain relation embeddings from a pre-trained language model, we encode word pairs using a (manually or automatically generated) prompt, and we fine-tune the language model such that relationally similar word pairs yield similar output vectors. We find that the resulting relation embeddings are highly competitive on analogy (unsupervised) and relation classification (supervised) benchmarks, even without any task-specific fine-tuning. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ca60672-9710-48d9-aa91-4d2d8ee8f884Cited by top-tier papers8
- StoryAnalogy: Deriving Story-level Analogies from Large Language Models to Unlock Analogical UnderstandingCheng Jiayang, Lin Qiu, Tsz Ho Chan, Tianqing Fang et al.EMNLP 2023 · 8 citations
- No clues good clues: out of context Lexical Relation ClassificationLucia Pitarch, Jordi Bernad, Lacramioara Dranca, Carlos Bobed Lisbona et al.ACL 2023 · 4 citations
- Can language models learn analogical reasoning? Investigating training objectives and comparisons to human performanceMolly R. Petersen, Lonneke van der PlasEMNLP 2023 · 3 citations
- Learning Dynamic Contextualised Word Embeddings via Template-based Temporal AdaptationXiaohang Tang, Yi Zhou, Danushka BollegalaACL 2023 · 2 citations
- How Sememic Components Can Benefit Link Prediction for Lexico-Semantic Knowledge Graphs?Hansi Wang, Yue Wang, Qiliang Liang, Yang LiuEMNLP 2025
Builds on9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated PromptsTaylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace et al.EMNLP 2020 · 1,162 citations
- Inducing Relational Knowledge from BERTZied Bouraoui, José Camacho-Collados, Steven SchockaertAAAI 2020 · 183 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
Related papers
- Solving Hard Analogy Questions with Relation Embedding ChainsNitesh Kumar, Steven SchockaertEMNLP 2023
- Introducing Graph Context into Language Models through Parameter-Efficient Fine-Tuning for Lexical Relation MiningJingwen Sun, Zhiyi Tian, Yu He, Jingwei Sun et al.ACL 2025
- Modeling Complex Semantics Relation with Contrastively Fine-Tuned Relational EncodersNaïm Es-Sebbani, Esteban Marquer, Zied BouraouiACL 2025
- IELM: An Open Information Extraction Benchmark for Pre-Trained Language ModelsChenguang Wang, Xiao Liu, Dawn SongEMNLP 2022 · 3 citations
- When Phrases Meet Probabilities: Enabling Open Relation Extraction with Cooperating Large Language ModelsJiaxin Wang, Lingling Zhang, Wee Sun Lee, Yujie Zhong et al.ACL 2024
