Assessing LLMs for Serendipity Discovery in Knowledge Graphs: A Case for Drug Repurposing
Mengying Wang, Chenhui Ma, Ao Jiao, Tuo Liang, Pengjun Lu, Shrinidhi Hegde, Yu Yin, Evren Gurkan-Cavusoglu, Yinghui Wu
Abstract
Large Language Models (LLMs) have greatly advanced knowledge graph question answering (KGQA), yet existing systems are typically optimized for returning highly relevant but predictable answers. A missing yet desired capacity is to exploit LLMs to suggest surprise and novel ("serendipitious") answers. In this paper, we formally define the serendipity-aware KGQA task and propose the SerenQA framework to evaluate LLMs' ability to uncover unexpected insights in scientific KGQA tasks. SerenQA includes a rigorous serendipity metric based on relevance, novelty, and surprise, along with an expert-annotated benchmark derived from the Clinical Knowledge Graph for drug repurposing. Additionally, it features a structured evaluation pipeline encompassing three subtasks: knowledge retrieval, subgraph reasoning, and serendipity exploration. Our experiments reveal that while state-of-the-art LLMs perform well on retrieval, they still struggle to identify genuinely surprising and valuable discoveries, underscoring a significant room for future research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 11ef4f34-2bc3-434a-93df-41d1889e0910Builds on1
Related papers
- Contextualizing biological perturbation experiments through languageMenghua Wu, Russell Littman, Jacob Levine, Lin Qiu et al.ICLR 2025
- K-Paths: Reasoning over Graph Paths for Drug Repurposing and Drug Interaction PredictionTassallah Abdullahi, Ioanna Gemou, Nihal V. Nayak, Ghulam Murtaza et al.KDD 2025 · 2 citations
- SSR: Structured Subgraph Retrieval for Temporal Knowledge Graph Question Answering with LLMsYing Zhang, Li Zhang, Wenya Guo, Shilong Ping et al.SIGIR 2026
- ProgRAG: Hallucination-Resistant Progressive Retrieval and Reasoning over Knowledge GraphsMinbae Park, Hyemin Yang, Jeonghyun Kim, Kunsoo Park et al.AAAI 2026
- LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language ModelsParshin Shojaee, Ngoc-Hieu Nguyen, Kazem Meidani, Amir Barati Farimani et al.ICML 2025
