Fast and Accurate Deep Bidirectional Language Representations for Unsupervised Learning
Joongbo Shin, Yoonhyung Lee, Seunghyun Yoon, Kyomin Jung
Abstract
Even though BERT has achieved successful performance improvements in various supervised learning tasks, BERT is still limited by repetitive inferences on unsupervised tasks for the computation of contextual language representations. To resolve this limitation, we propose a novel deep bidirectional language model called a Transformer-based Text Autoencoder (T-TA). The T-TA computes contextual language representations without repetition and displays the benefits of a deep bidirectional architecture, such as that of BERT. In computation time experiments in a CPU environment, the proposed T-TA performs over six times faster than the BERT-like model on a reranking task and twelve times faster on a semantic similarity task. Furthermore, the T-TA shows competitive or even better accuracies than those of BERT on the above tasks. Code is available at https://github.com/joongbo/tta .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dbc7264d-c1bb-42fc-a793-713420221ad3Cited by top-tier papers2
- Transcormer: Transformer for Sentence Scoring with Sliding Language ModelingKaitao Song, Yichong Leng, Xu Tan, Yicheng Zou et al.NeurIPS 2022 · 12 citations
- ExLM: Rethinking the Impact of [MASK] Tokens in Masked Language ModelsKangjie Zheng, Junwei Yang, Siyue Liang, Bin Feng et al.ICML 2025
Builds on1
Related papers
- Repetition Improves Language Model EmbeddingsJacob Mitchell Springer, Suhas Kotha, Daniel Fried, Graham Neubig et al.ICLR 2025
- Lexical Simplification with Pretrained EncodersJipeng Qiang, Yun Li, Yi Zhu, Yunhao Yuan et al.AAAI 2020 · 86 citations
- Diffusion vs. Autoregressive Language Models: A Text Embedding PerspectiveSiyue Zhang, Yilun Zhao, Liyuan Geng, Arman Cohan et al.EMNLP 2025
- Bidirectional Variational Inference for Non-Autoregressive Text-to-SpeechYoonhyung Lee, Joongbo Shin, Kyomin JungICLR 2021 · 42 citations
- EEL: Efficiently Encoding Lattices for RerankingPrasann Singhal, Jiacheng Xu, Xi Ye, Greg DurrettACL 2023
