On the Theoretical Limitations of Embedding-based Link Prediction
Samy Badreddine, Emile van Krieken, Luciano Serafini
摘要
Neural networks often map low-dimensional embeddings to high-dimensional output spaces. Usually, the output layer is linear, which can create a rank bottleneck that limits the functions a model can represent. Such bottlenecks are ubiquitous in link prediction models, such as knowledge graph embeddings (KGEs), as the output space of entities can be orders of magnitude larger than the embedding dimension. We investigate how rank bottlenecks limit model expressivity for fitting the training data. While previous work focused on sufficient bounds on the embedding dimension required for specific KGEs, we show necessary bounds for all KGEs with a linear output layer, which grow with graph size and connectivity. We also consider a non-linear output layer using mixtures to break the bottleneck without significant parameter overhead. Empirically, we show that models using this non-linear layer improve in ranking performance and probabilistic fit for large and dense datasets at a low parameter cost, as predicted by our theory. Our work reveals how linear output layers limit KGEs and motivates non-linear alternatives for scaling to large and dense graphs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Composition-based Multi-Relational Graph Convolutional NetworksShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Partha P. TalukdarICLR 2020 · 被引用 1,105 次
- BoxE: A Box Embedding Model for Knowledge Base CompletionRalph Abboud, Ismail Ilkan Ceylan, Thomas Lukasiewicz, Tommaso SalvatoriNeurIPS 2020 · 被引用 245 次
- You CAN Teach an Old Dog New Tricks! On Training Knowledge Graph EmbeddingsDaniel Ruffinelli, Samuel Broscheit, Rainer GemullaICLR 2020 · 被引用 238 次
- Node Embeddings and Exact Low-Rank Representations of Complex NetworksSudhanshu Chanpuriya, Cameron Musco, Konstantinos Sotiropoulos, Charalampos E. TsourakakisNeurIPS 2020 · 被引用 41 次
相关 Paper
- On the Softmax Bottleneck of Recurrent Language ModelsDwarak Govind Parthiban, Yongyi Mao, Diana InkpenAAAI 2021 · 被引用 3 次
- ParamE: Regarding Neural Network Parameters as Relation Embeddings for Knowledge Graph CompletionFeihu Che, Dawei Zhang, Jianhua Tao, Mingyue Niu 等AAAI 2020 · 被引用 54 次
- Contextual Parameter Generation for Knowledge Graph Link PredictionGeorge Stoica, Otilia Stretcu, Emmanouil Antonios Platanios, Tom M. Mitchell 等AAAI 2020 · 被引用 46 次
- Interpreting Knowledge Graph Relation Representation from Word EmbeddingsCarl Allen, Ivana Balazevic, Timothy M. HospedalesICLR 2021 · 被引用 7 次
- LowFER: Low-rank Bilinear Pooling for Link PredictionSaadullah Amin, Stalin Varanasi, Katherine Ann Dunfield, Günter NeumannICML 2020 · 被引用 43 次
