Unsupervised Cross-Task Generalization via Retrieval Augmentation
Bill Yuchen Lin, Kangmin Tan, Chris Miller, Beiwen Tian, Xiang Ren
Abstract
Humans can perform unseen tasks by recalling relevant skills acquired previously and then generalizing them to the target tasks, even if there is no supervision at all. In this paper, we aim to improve this kind of cross-task generalization ability of massive multi-task language models, such as T0 and FLAN, in an unsupervised setting. We propose a retrieval-augmentation method named ReCross that takes a few unlabelled examples as queries to retrieve a small subset of upstream data and uses them to update the multi-task model for better generalization. ReCross is a straightforward yet effective retrieval method that combines both efficient dense retrieval and effective pair-wise reranking. Our results and analysis show that it significantly outperforms both non-retrieval methods and other baseline methods. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context LearningBill Yuchen Lin, Abhilasha Ravichander, Ximing Lu, Nouha Dziri et al.ICLR 2024 · 299 citations
- Exploring the Benefits of Training Expert Language Models over Instruction TuningJoel Jang, Seungone Kim, Seonghyeon Ye, Doyoung Kim et al.ICML 2023 · 97 citations
- Improving Few-Shot Generalization by Exploring and Exploiting Auxiliary DataAlon Albalak, Colin A. Raffel, William Yang WangNeurIPS 2023 · 17 citations
- Cappy: Outperforming and Boosting Large Multi-Task LMs with a Small ScorerBowen Tan, Yun Zhu, Lijuan Liu, Eric P. Xing et al.NeurIPS 2023 · 11 citations
- Guess the Instruction! Flipped Learning Makes Language Models Stronger Zero-Shot LearnersSeonghyeon Ye, Doyoung Kim, Joel Jang, Joongbo Shin et al.ICLR 2023 · 10 citations
Builds on10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
Related papers
- Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-InZichun Yu, Chenyan Xiong, Shi Yu, Zhiyuan LiuACL 2023 · 15 citations
- A Multi-Task Embedder For Retrieval Augmented LLMsPeitian Zhang, Zheng Liu, Shitao Xiao, Zhicheng Dou et al.ACL 2024
- Augmenting Zero-Shot Dense Retrievers with Plug-in Mixture-of-MemoriesSuyu Ge, Chenyan Xiong, Corby Rosset, Arnold Overwijk et al.EMNLP 2023 · 4 citations
- CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLPQinyuan Ye, Bill Yuchen Lin, Xiang RenEMNLP 2021 · 103 citations
- Improving Passage Retrieval with Zero-Shot Question GenerationDevendra Singh Sachan, Mike Lewis, Mandar Joshi, Armen Aghajanyan et al.EMNLP 2022 · 69 citations
