FiTs: Fine-Grained Two-Stage Training for Knowledge-Aware Question Answering
Qichen Ye, Bowen Cao, Nuo Chen, Weiyuan Xu, Yuexian Zou
Abstract
Knowledge-aware question answering (KAQA) requires the model to answer questions over a knowledge base, which is essential for both open-domain QA and domain-specific QA, especially when language models alone cannot provide all the knowledge needed. Despite the promising result of recent KAQA systems which tend to integrate linguistic knowledge from pre-trained language models (PLM) and factual knowledge from knowledge graphs (KG) to answer complex questions, a bottleneck exists in effectively fusing the representations from PLMs and KGs because of (i) the semantic and distributional gaps between them, and (ii) the difficulties in joint reasoning over the provided knowledge from both modalities. To address the above two problems, we propose a Fine-grained Two-stage training framework (FiTs) to boost the KAQA system performance: The first stage aims at aligning representations from the PLM and the KG, thus bridging the modality gaps between them, named knowledge adaptive post-training. The second stage, called knowledge-aware fine-tuning, aims to improve the model's joint reasoning ability based on the aligned representations. In detail, we fine-tune the post-trained model via two auxiliary self-supervised tasks in addition to the QA supervision. Extensive experiments demonstrate that our approach achieves state-of-the-art performance on three benchmarks in the commonsense reasoning (i.e., CommonsenseQA, OpenbookQA) and medical question answering (i.e., MedQA-USMILE) domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 482bee52-6109-4544-9eb4-a13e48dc0026Cited by top-tier papers4
- Retrieval is Accurate GenerationBowen Cao, Deng Cai, Leyang Cui, Xuxin Cheng et al.ICLR 2024 · 13 citations
- Graph Reasoning Transformers for Knowledge-Aware Question AnsweringRuilin Zhao, Feng Zhao, Liang Hu, Guandong XuAAAI 2024 · 10 citations
- KGE Calibrator: An Efficient Probability Calibration Method of Knowledge Graph Embedding Models for Trustworthy Link PredictionYang Yang, Mohan Timilsina, Edward CurryEMNLP 2025
- Video-Text as Game Players: Hierarchical Banzhaf Interaction for Cross-Modal Representation LearningPeng Jin, Jinfa Huang, Pengfei Xiong, Shangxuan Tian et al.CVPR 2023
Builds on6
- Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation LearningWeixin Liang, Yuhui Zhang, Yongchan Kwon, Serena Yeung et al.NeurIPS 2022 · 834 citations
- Beyond I.I.D.: Three Levels of Generalization for Question Answering on Knowledge BasesYu Gu, Sue Kase, Michelle Vanni, Brian M. Sadler et al.WWW 2021 · 304 citations
- Graph-Based Reasoning over Heterogeneous External Knowledge for Commonsense Question AnsweringShangwen Lv, Daya Guo, Jingjing Xu, Duyu Tang et al.AAAI 2020 · 224 citations
- SPARQA: Skeleton-Based Semantic Parsing for Complex Questions over Knowledge BasesYawei Sun, Lingling Zhang, Gong Cheng, Yuzhong QuAAAI 2020 · 143 citations
- Expectation-Maximization Contrastive Learning for Compact Video-and-Language RepresentationsPeng Jin, Jinfa Huang, Fenglin Liu, Xian Wu et al.NeurIPS 2022 · 105 citations
Related papers
- Relation-Aware Language-Graph Transformer for Question AnsweringJinyoung Park, Hyeong Kyu Choi, Juyeon Ko, Hyeon-Jin Park et al.AAAI 2023 · 16 citations
- ReasoningLM: Enabling Structural Subgraph Reasoning in Pre-trained Language Models for Question Answering over Knowledge GraphJinhao Jiang, Kun Zhou, Wayne Xin Zhao, Yaliang Li et al.EMNLP 2023 · 26 citations
- Modality-Aware Integration with Large Language Models for Knowledge-Based Visual Question AnsweringJunnan Dong, Qinggang Zhang, Huachi Zhou, Daochen Zha et al.ACL 2024 · 11 citations
- GreaseLM: Graph REASoning Enhanced Language ModelsXikun Zhang, Antoine Bosselut, Michihiro Yasunaga, Hongyu Ren et al.ICLR 2022 · 285 citations
- A Knowledge-Injected Curriculum Pretraining Framework for Question AnsweringXin Lin, Tianhuang Su, Zhenya Huang, Shangzi Xue et al.WWW 2024 · 3 citations
