Active Learning on Pre-trained Language Model with Task-Independent Triplet Loss
Seungmin Seo, Donghyun Kim, Youbin Ahn, Kyong-Ho Lee
Abstract
Active learning attempts to maximize a task model’s performance gain by obtaining a set of informative samples from an unlabeled data pool. Previous active learning methods usually rely on specific network architectures or task-dependent sample acquisition algorithms. Moreover, when selecting a batch sample, previous works suffer from insufficient diversity of batch samples because they only consider the informativeness of each sample. This paper proposes a task-independent batch acquisition method using triplet loss to distinguish hard samples in an unlabeled data pool with similar features but difficult to identify labels. To assess the effectiveness of the proposed method, we compare the proposed method with state-of-the-art active learning methods on two tasks, relation extraction and sentence classification. Experimental results show that our method outperforms baselines on the benchmark datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d0f5c07-ce6a-4f3e-af98-e5f3d2f10c14Cited by top-tier papers5
- A Survey of Active Learning for Natural Language ProcessingZhisong Zhang, Emma Strubell, Eduard H. HovyEMNLP 2022 · 60 citations
- Inconsistency-Based Data-Centric Active Open-Set AnnotationRuiyu Mao, Ouyang Xu, Yunhui GuoAAAI 2024 · 7 citations
- CoLAL: Co-learning Active Learning for Text ClassificationLinh Le, Genghong Zhao, Xia Zhang, Guido Zuccon et al.AAAI 2024 · 5 citations
- Heterogeneous Adversarial Play in Interactive EnvironmentsManjie Xu, Xinyi Yang, Jiayu Zhan, Wei Liang et al.NeurIPS 2025 · 4 citations
- TriG-NER: Triplet-Grid Framework for Discontinuous Named Entity RecognitionRina Carines Cabral, Soyeon Caren Han, Areej Alhassan, Riza Batista-Navarro et al.WWW 2025 · 2 citations
Builds on7
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 662 citations
- Active Learning for BERT: An Empirical StudyLiat Ein-Dor, Alon Halfon, Ariel Gera, Eyal Shnarch et al.EMNLP 2020 · 144 citations
- Cold-start Active Learning through Self-supervised Language ModelingMichelle Yuan, Hsuan-Tien Lin, Jordan L. Boyd-GraberEMNLP 2020 · 128 citations
- Effective Unsupervised Domain Adaptation with Adversarially Trained Language ModelsThuy-Trang Vu, Dinh Phung, Gholamreza HaffariEMNLP 2020 · 19 citations
Related papers
- The Dilemma of TriHard Loss and an Element-Weighted TriHard Loss for Person Re-IdentificationYihao Lv, Youzhi Gu, Xinggao LiuNeurIPS 2020 · 11 citations
- Active Learning by Acquiring Contrastive ExamplesKaterina Margatina, Giorgos Vernikos, Loïc Barrault, Nikolaos AletrasEMNLP 2021 · 8 citations
- Influence Selection for Active LearningZhuoming Liu, Hao Ding, Huaping Zhong, Weijia Li et al.ICCV 2021 · 125 citations
- State-Relabeling Adversarial Active LearningBeichen Zhang, Liang Li, Shijie Yang, Shuhui Wang et al.CVPR 2020
- Task-Aware Variational Adversarial Active LearningKwanyoung Kim, Dongwon Park, Kwang In Kim, Se Young ChunCVPR 2021
