Prompt-Based Rule Discovery and Boosting for Interactive Weakly-Supervised Learning
Rongzhi Zhang, Yue Yu, Pranav Shetty, Le Song, Chao Zhang
摘要
Weakly-supervised learning (WSL) has shown promising results in addressing label scarcity on many NLP tasks, but manually designing a comprehensive, high-quality labeling rule set is tedious and difficult. We study interactive weakly-supervised learning-the problem of iteratively and automatically discovering novel labeling rules from data to improve the WSL model. Our proposed model, named PR-BOOST, achieves this goal via iterative promptbased rule discovery and model boosting. It uses boosting to identify large-error instances and then discovers candidate rules from them by prompting pre-trained LMs with rule templates. The candidate rules are judged by human experts, and the accepted rules are used to generate complementary weak labels and strengthen the current model. Experiments on four tasks show PRBOOST outperforms state-of-the-art WSL baselines up to 7.1%, and bridges the gaps with fully supervised models.Our Implementation is available at https: //github.com/rz-zhang/PRBoost .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- A Survey of Active Learning for Natural Language ProcessingZhisong Zhang, Emma Strubell, Eduard H. HovyEMNLP 2022 · 被引用 60 次
- Towards Interactivity and Interpretability: A Rationale-based Legal Judgment Prediction FrameworkYiquan Wu, Yifei Liu, Weiming Lu, Yating Zhang 等EMNLP 2022 · 被引用 33 次
- BLESS: Benchmarking Large Language Models on Sentence SimplificationTannon Kew, Alison Chi, Laura Vásquez-Rodríguez, Sweta Agrawal 等EMNLP 2023 · 被引用 15 次
- Cold-Start Data Selection for Better Few-shot Language Model Fine-tuning: A Prompt-based Uncertainty Propagation ApproachYue Yu, Rongzhi Zhang, Ran Xu, Jieyu Zhang 等ACL 2023 · 被引用 12 次
- De-biased Attention Supervision for Text Classification with CausalityYiquan Wu, Yifei Liu, Ziyu Zhao, Weiming Lu 等AAAI 2024 · 被引用 10 次
它引用的顶会 Paper19
- True Few-Shot Learning with Language ModelsEthan Perez, Douwe Kiela, Kyunghyun ChoNeurIPS 2021 · 被引用 547 次
- Text Classification Using Label Names Only: A Language Model Self-Training ApproachYu Meng, Yunyi Zhang, Jiaxin Huang, Chenyan Xiong 等EMNLP 2020 · 被引用 203 次
- Active Learning for BERT: An Empirical StudyLiat Ein-Dor, Alon Halfon, Ariel Gera, Eyal Shnarch 等EMNLP 2020 · 被引用 144 次
- Contextualized Weak Supervision for Text ClassificationDheeraj Mekala, Jingbo ShangACL 2020 · 被引用 121 次
- BOND: BERT-Assisted Open-Domain Named Entity Recognition with Distant SupervisionChen Liang, Yue Yu, Haoming Jiang, Siawpeng Er 等KDD 2020 · 被引用 118 次
相关 Paper
- RulePrompt: Weakly Supervised Text Classification with Prompting PLMs and Self-Iterative Logical RulesMiaomiao Li, Jiaqi Zhu, Yang Wang, Yi Yang 等WWW 2024 · 被引用 5 次
- Adaptive Rule Discovery for Labeling Text DataSainyam Galhotra, Behzad Golshan, Wang-Chiew TanSIGMOD 2021 · 被引用 14 次
- Interactive Weak Supervision: Learning Useful Heuristics for Data LabelingBenedikt Boecking, Willie Neiswanger, Eric P. Xing, Artur DubrawskiICLR 2021 · 被引用 8 次
- Mitigating Data Scarcity in Supervised Machine Learning Through Reinforcement Learning Guided Data GenerationChengliang Chai, Kaisen Jin, Nan Tang, Ju Fan 等ICDE 2024 · 被引用 7 次
- Learning from True-False Labels via Multi-modal Prompt RetrievingZhongnian Li, Jinghao Xu, Peng Ying, Meng Wei 等ICML 2025
