Leveraging Estimated Transferability Over Human Intuition for Model Selection in Text Ranking
Jun Bai, Zhuofan Chen, Zhenzi Li, Hanhua Hong, Jianfei Zhang, Chen Li, Chenghua Lin, Wenge Rong
摘要
Text ranking has witnessed significant advancements, attributed to the utilization of dual-encoder enhanced by Pre-trained Language Models (PLMs). Given the proliferation of available PLMs, selecting the most effective one for a given dataset has become a non-trivial challenge. As a promising alternative to human intuition and brute-force fine-tuning, Transferability Estimation (TE) has emerged as an effective approach to model selection. However, current TE methods are primarily designed for classification tasks, and their estimated transferability may not align well with the objectives of text ranking. To address this challenge, we propose to compute the expected rank as transferability, explicitly reflecting the model’s ranking capability. Furthermore, to mitigate anisotropy and incorporate training dynamics, we adaptively scale isotropic sentence embeddings to yield an accurate expected rank score. Our resulting method, Adaptive Ranking Transferability (AiRTran), can effectively capture subtle differences between models. On challenging model selection scenarios across various text ranking datasets, it demonstrates significant improvements over previous classification-oriented TE methods, human intuition, and ChatGPT with minor time consumption.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- On the Sentence Embeddings from Pre-trained Language ModelsBohan Li, Hao Zhou, Junxian He, Mingxuan Wang 等EMNLP 2020 · 被引用 538 次
- LogME: Practical Assessment of Pre-trained Models for Transfer LearningKaichao You, Yong Liu, Jianmin Wang, Mingsheng LongICML 2021 · 被引用 253 次
- MuTual: A Dataset for Multi-Turn Dialogue ReasoningLeyang Cui, Yu Wu, Shujie Liu, Yue Zhang 等ACL 2020 · 被引用 115 次
相关 Paper
- ETran: Energy-Based Transferability EstimationMohsen Gholami, Mohammad Akbari, Xinglu Wang, Behnam Kamranian 等ICCV 2023 · 被引用 21 次
- PEFTDiff: Diffusion-Guided Transferability Estimation for Parameter-Efficient Fine-TuningPrafful Kumar Khoba, Zijian Wang, Chetan Arora, Mahsa BaktashmotlaghICCV 2025 · 被引用 2 次
- How NOT to benchmark your SITE metric: Beyond Static Leaderboards and Towards Realistic Evaluation.Prabhant Singh, Sibylle Hess, Joaquin VanschorenICLR 2026 · 被引用 2 次
- Transcormer: Transformer for Sentence Scoring with Sliding Language ModelingKaitao Song, Yichong Leng, Xu Tan, Yicheng Zou 等NeurIPS 2022 · 被引用 12 次
- Incorporating Explicit Knowledge in Pre-trained Language Models for Passage Re-rankingQian Dong, Yiding Liu, Suqi Cheng, Shuaiqiang Wang 等SIGIR 2022 · 被引用 14 次
