OTKGE: Multi-modal Knowledge Graph Embeddings via Optimal Transport
Zongsheng Cao, Qianqian Xu, Zhiyong Yang, Yuan He, Xiaochun Cao, Qingming Huang
摘要
Multi-modal knowledge graph embeddings (KGE) have caught more and more attention in learning representations of entities and relations for link prediction tasks. Different from previous uni-modal KGE approaches, multi-modal KGE can leverage expressive knowledge from a wealth of modalities (image, text, etc.), leading to more comprehensive representations of real-world entities. However, the critical challenge along this course lies in that the multi-modal embedding spaces are usually heterogeneous. In this sense, direct fusion will destroy the inherent spatial structure of different modal embeddings. To overcome this challenge, we revisit multi-modal KGE from a distributional alignment perspective and propose optimal transport knowledge graph embeddings (OTKGE). Specifically, we model the multi-modal fusion procedure as a transport plan moving different modal embeddings to a unified space by minimizing the Wasserstein distance between multi-modal distributions. Theoretically, we show that by minimizing the Wasserstein distance between the individual modalities and the unified embedding space, the final results are guaranteed to maintain consistency and comprehensiveness. Moreover, experimental results on well-established multi-modal knowledge graph completion benchmarks show that our OTKGE achieves state-of-the-art performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Multimodal Optimal Transport-based Co-Attention Transformer with Global Structure Consistency for Survival PredictionYingxue Xu, Hao ChenICCV 2023 · 被引用 132 次
- Multimodal Prototyping for cancer survival predictionAndrew H. Song, Richard J. Chen, Guillaume Jaume, Anurag J. Vaidya 等ICML 2024 · 被引用 53 次
- LAFA: Multimodal Knowledge Graph Completion with Link Aware Fusion and AggregationBin Shang, Yinliang Zhao, Jun Liu, Di WangAAAI 2024 · 被引用 40 次
- NativE: Multi-modal Knowledge Graph Completion in the WildYichi Zhang, Zhuo Chen, Lingbing Guo, Yajing Xu 等SIGIR 2024 · 被引用 39 次
- Contrastive Intra- and Inter-Modality Generation for Enhancing Incomplete Multimedia RecommendationZhenghong Lin, Yanchao Tan, Yunfei Zhan, Weiming Liu 等ACM MM 2023 · 被引用 27 次
它引用的顶会 Paper4
- Model Fusion via Optimal TransportSidak Pal Singh, Martin JaggiNeurIPS 2020 · 被引用 330 次
- Dual Quaternion Knowledge Graph EmbeddingsZongsheng Cao, Qianqian Xu, Zhiyong Yang, Xiaochun Cao 等AAAI 2021 · 被引用 186 次
- Low-Dimensional Hyperbolic Knowledge Graph EmbeddingsInes Chami, Adva Wolf, Da-Cheng Juan, Frederic Sala 等ACL 2020 · 被引用 48 次
- Multi-modal Multi-relational Feature Aggregation Network for Medical Knowledge Representation LearningYingying Zhang, Quan Fang, Shengsheng Qian, Changsheng XuACM MM 2020 · 被引用 20 次
相关 Paper
- Bridging the Space Gap: Unifying Geometry Knowledge Graph Embedding with Optimal TransportYuhan Liu, Zelin Cao, Xing Gao, Ji Zhang 等WWW 2024 · 被引用 11 次
- VL-KGE: Vision-Language Models Meet Knowledge Graph EmbeddingsAthanasios Efthymiou, Stevan Rudinac, Monika Kackovic, Nachoem Wijnberg 等WWW 2026 · 被引用 2 次
- REMOTE: A Unified Multimodal Relation Extraction Framework with Multilevel Optimal Transport and Mixture-of-ExpertsXinkui Lin, Yongxiu Xu, Minghao Tang, Shilong Zhang 等ACM MM 2025 · 被引用 2 次
- MoSE: Modality Split and Ensemble for Multimodal Knowledge Graph CompletionYu Zhao, Xiangrui Cai, Yike Wu, Haiwei Zhang 等EMNLP 2022 · 被引用 66 次
- Multiple Heads are Better than One: Mixture of Modality Knowledge Experts for Entity Representation LearningYichi Zhang, Zhuo Chen, Lingbing Guo, Yajing Xu 等ICLR 2025 · 被引用 1 次
