Cross-domain Knowledge Distillation for Retrieval-based Question Answering Systems
Cen Chen, Chengyu Wang, Minghui Qiu, Dehong Gao, Linbo Jin, Wang Li
摘要
Question Answering (QA) systems have been extensively studied in both academia and the research community due to their wide real-world applications. When building such industrial-scale QA applications, we are facing two prominent challenges, i.e., i) lacking a sufficient amount of training data to learn an accurate model and ii) requiring high inference speed for online model serving. There are generally two ways to mitigate the above-mentioned problems. One is to adopt transfer learning to leverage information from other domains; the other is to distill the "dark knowledge" from a large teacher model to small student models. The former usually employs parameter sharing mechanisms for knowledge transfer, but does not utilize the "dark knowledge" of pre-trained large models. The latter usually does not consider the cross-domain information from other domains. We argue that these two types of methods can be complementary to each other. Hence in this work, we provide a new perspective on the potential of the teacher-student paradigm facilitating cross-domain transfer learning, where the teacher and student tasks belong to heterogeneous domains, with the goal to improve the student model's performance in the target domain. Our framework considers the "dark knowledge" learned from large teacher models and also leverages the adaptive hints to alleviate the domain differences between teacher and student models. Extensive experiments have been conducted on two text matching tasks for retrieval-based QA systems. Results show the proposed method has better performance than the competing methods including the existing state-of-the-art transfer learning methods. We have also deployed our method in an online production system and observed significant improvements compared to the existing approaches in terms of both accuracy and cross-domain robustness. CCS CONCEPTS • Applied computing → Electronic commerce; • Information systems → Retrieval models and ranking.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- TMLKD: Few-shot Trajectory Metric Learning via Knowledge DistillationDanling Lai, Jiajie Xu, Jianfeng Qu, Pingfu Chao 等VLDB 2025 · 被引用 1 次
- Meta-KD: A Meta Knowledge Distillation Framework for Language Model Compression across DomainsHaojie Pan, Chengyu Wang, Minghui Qiu, Yichang Zhang 等ACL 2021
它引用的顶会 Paper2
相关 Paper
- TANDA: Transfer and Adapt Pre-Trained Transformer Models for Answer Sentence SelectionSiddhant Garg, Thuy Vu, Alessandro MoschittiAAAI 2020 · 被引用 229 次
- DoQA - Accessing Domain-Specific FAQs via Conversational QAJon Ander Campos, Arantxa Otegi, Aitor Soroa, Jan Deriu 等ACL 2020 · 被引用 2 次
- Heterogeneous Continual LearningDivyam Madaan, Hongxu Yin, Wonmin Byeon, Jan Kautz 等CVPR 2023
- Collaborative Enhancement of Large and Small Models for Question Answering via Dual Knowledge TransferShaofei Wang, Yunan Liu, Xiaolan Tang, Wenlong ChenAAAI 2026
- Learning Systems Expansion with Efficient Heterogeneity-aware Knowledge TransferGaole Dai, Huatao Xu, Yifan Yang, Rui Tan 等AAAI 2026
