B-PROP: Bootstrapped Pre-training with Representative Words Prediction for Ad-hoc Retrieval
Xinyu Ma, Jiafeng Guo, Ruqing Zhang, Yixing Fan, Yingyan Li, Xueqi Cheng
摘要
Pre-training and fine-tuning have achieved remarkable success in many downstream natural language processing (NLP) tasks. Recently, pre-training methods tailored for information retrieval (IR) have also been explored, and the latest success is the PROP method which has reached new SOTA on a variety of ad-hoc retrieval benchmarks. The basic idea of PROP is to construct the representative words prediction (ROP) task for pre-training inspired by the query likelihood model. Despite its exciting performance, the effectiveness of PROP might be bounded by the classical unigram language model adopted in the ROP task construction process. To tackle this problem, we propose a bootstrapped pre-training method (namely B-PROP) based on BERT for ad-hoc retrieval. The key idea is to use the powerful contextual language model BERT to replace the classical unigram language model for the ROP task construction, and re-train BERT itself towards the tailored objective for IR. Specifically, we introduce a novel contrastive method, inspired by the divergence-from-randomness idea, to leverage BERT's self-attention mechanism to sample representative words from the document. By further fine-tuning on downstream ad-hoc retrieval tasks, our method achieves significant improvements over PROP and other baselines, and further pushes forward the SOTA on a variety of ad-hoc retrieval tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Fine-Grained Distillation for Long Document RetrievalYucheng Zhou, Tao Shen, Xiubo Geng, Chongyang Tao 等AAAI 2024 · 被引用 44 次
- Pre-train a Discriminative Text Encoder for Dense Retrieval via Contrastive Span PredictionXinyu Ma, Jiafeng Guo, Ruqing Zhang, Yixing Fan 等SIGIR 2022 · 被引用 33 次
- Webformer: Pre-training with Web Pages for Information RetrievalYu Guo, Zhengyi Ma, Jiaxin Mao, Hongjin Qian 等SIGIR 2022 · 被引用 30 次
- Axiomatically Regularized Pre-training for Ad hoc SearchJia Chen, Yiqun Liu, Yan Fang, Jiaxin Mao 等SIGIR 2022 · 被引用 19 次
- Generative Retrieval Meets Multi-Graded RelevanceYubao Tang, Ruqing Zhang, Jiafeng Guo, Maarten de Rijke 等NeurIPS 2024 · 被引用 18 次
它引用的顶会 Paper5
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 被引用 2,453 次
- Pre-training Tasks for Embedding-based Large-scale RetrievalWei-Cheng Chang, Felix X. Yu, Yin-Wen Chang, Yiming Yang 等ICLR 2020 · 被引用 325 次
- SentiLARE: Sentiment-Aware Language Representation Learning with Linguistic KnowledgePei Ke, Haozhe Ji, Siyang Liu, Xiaoyan Zhu 等EMNLP 2020 · 被引用 138 次
- A Linguistic Study on Relevance Modeling in Information RetrievalYixing Fan, Jiafeng Guo, Xinyu Ma, Ruqing Zhang 等WWW 2021 · 被引用 15 次
相关 Paper
- COCO-DR: Combating the Distribution Shift in Zero-Shot Dense Retrieval with Contrastive and Distributionally Robust LearningYue Yu, Chenyan Xiong, Si Sun, Chao Zhang 等EMNLP 2022 · 被引用 21 次
- Table Search Using a Deep Contextualized Language ModelZhiyu Chen, Mohamed Trabelsi, Jeff Heflin, Yinan Xu 等SIGIR 2020 · 被引用 48 次
- H-ERNIE: A Multi-Granularity Pre-Trained Language Model for Web SearchXiaokai Chu, Jiashu Zhao, Lixin Zou, Dawei YinSIGIR 2022 · 被引用 11 次
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERTOmar Khattab, Matei ZahariaSIGIR 2020 · 被引用 1,246 次
- Cross-lingual Language Model Pretraining for RetrievalPuxuan Yu, Hongliang Fei, Ping LiWWW 2021 · 被引用 42 次
