Which *BERT? A Survey Organizing Contextualized Encoders
Patrick Xia, Shijie Wu, Benjamin Van Durme
Abstract
Pretrained contextualized text encoders are now a staple of the NLP community. We present a survey on language representation learning with the aim of consolidating a series of shared lessons learned across a variety of recent efforts. While significant advancements continue at a rapid pace, we find that enough has now been discovered, in different directions, that we can begin to organize advances according to common themes. Through this organization, we highlight important considerations when interpreting recent contributions and choosing which model to use.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 605f27bb-adac-4d6a-9a40-ebe603baa595Cited by top-tier papers6
- The Surprising Performance of Simple Baselines for Misinformation DetectionKellin Pelrine, Jacob Danovitch, Reihaneh RabbanyWWW 2021 · 79 citations
- When Hearst Is not Enough: Improving Hypernymy Detection from Corpus with Distributional ModelsChanglong Yu, Jialong Han, Peifeng Wang, Yangqiu Song et al.EMNLP 2020 · 14 citations
- Textual Manifold-based Defense Against Natural Language Adversarial ExamplesDang Minh Nguyen, Anh Tuan LuuEMNLP 2022 · 13 citations
- Structure and Semantics Preserving Document RepresentationsNatraj Raman, Sameena Shah, Manuela VelosoSIGIR 2022 · 5 citations
- A Predictive Factor Analysis of Social Biases and Task-Performance in Pretrained Masked Language ModelsYi Zhou, José Camacho-Collados, Danushka BollegalaEMNLP 2023 · 1 citation
Builds on32
- Towards Evaluating the Robustness of Neural NetworksNicholas Carlini, David A. WagnerS&P 2017 · 9,786 citations
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
- MPNet: Masked and Permuted Pre-training for Language UnderstandingKaitao Song, Xu Tan, Tao Qin, Jianfeng Lu et al.NeurIPS 2020 · 1,957 citations
- VL-BERT: Pre-training of Generic Visual-Linguistic RepresentationsWeijie Su, Xizhou Zhu, Yue Cao, Bin Li et al.ICLR 2020 · 1,825 citations
Related papers
- KnowLog: Knowledge Enhanced Pre-trained Language Model for Log UnderstandingLipeng Ma, Weidong Yang, Bo Xu, Sihang Jiang et al.ICSE 2024 · 26 citations
- TextNAS: A Neural Architecture Search Space Tailored for Text RepresentationYujing Wang, Yaming Yang, Yiren Chen, Jing Bai et al.AAAI 2020 · 66 citations
- Learning Visual Representations via Language-Guided SamplingMohamed El Banani, Karan Desai, Justin JohnsonCVPR 2023
- Coreferential Reasoning Learning for Language RepresentationDeming Ye, Yankai Lin, Jiaju Du, Zhenghao Liu et al.EMNLP 2020 · 164 citations
- Nugget: Neural Agglomerative Embeddings of TextGuanghui Qin, Benjamin Van DurmeICML 2023 · 24 citations
