Creating Training Sets via Weak Indirect Supervision
Jieyu Zhang, Bohan Wang, Xiangchen Song, Yujing Wang, Yaming Yang, Jing Bai, Alexander Ratner
摘要
Creating labeled training sets has become one of the major roadblocks in machine learning. To address this, recent Weak Supervision (WS) frameworks synthesize training labels from multiple potentially noisy supervision sources. However, existing frameworks are restricted to supervision sources that share the same output space as the target task. To extend the scope of usable sources, we formulate Weak Indirect Supervision (WIS), a new research problem for automatically synthesizing training labels based on indirect supervision sources that have different output label spaces. To overcome the challenge of mismatched output spaces, we develop a probabilistic modeling approach, PLRM, which uses user-provided label relations to model and leverage indirect supervision sources. Moreover, we provide a theoretically-principled test of the distinguishability of PLRM for unseen labels, along with a generalization bound. On both image and text classification tasks as well as an industrial advertising application, we demonstrate the advantages of PLRM by outperforming baselines by a margin of 2%-9%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- CoAnnotating: Uncertainty-Guided Work Allocation between Human and Large Language Models for Data AnnotationMinzhi Li, Taiwei Shi, Caleb Ziems, Min-Yen Kan 等EMNLP 2023 · 被引用 33 次
- Losses over Labels: Weakly Supervised Learning via Direct Loss ConstructionDylan Sam, J. Zico KolterAAAI 2023 · 被引用 14 次
- Understanding Programmatic Weak Supervision via Source-aware Influence FunctionJieyu Zhang, Haonan Wang, Cheng-Yu Hsieh, Alexander J. RatnerNeurIPS 2022 · 被引用 13 次
- Robust Weak Supervision with Variational Auto-EncodersFrancesco Tonolini, Nikolaos Aletras, Yunlong Jiao, Gabriella KazaiICML 2023 · 被引用 7 次
- WeShap: Weak Supervision Source Evaluation with Shapley ValuesNaiqing Guan, Nick KoudasVLDB 2025
它引用的顶会 Paper8
- Few-shot Relation Extraction via Bayesian Meta-learning on Relation GraphsMeng Qu, Tianyu Gao, Louis-Pascal A. C. Xhonneux, Jian TangICML 2020 · 被引用 131 次
- Fast and Three-rious: Speeding Up Weak Supervision with Triplet MethodsDaniel Y. Fu, Mayee F. Chen, Frederic Sala, Sarah M. Hooper 等ICML 2020 · 被引用 130 次
- Co-Tuning for Transfer LearningKaichao You, Zhi Kou, Mingsheng Long, Jianmin WangNeurIPS 2020 · 被引用 105 次
- Weakly Supervised Sequence Tagging from Noisy RulesEsteban Safranchik, Shiying Luo, Stephen H. BachAAAI 2020 · 被引用 90 次
- Adversarial Multi Class Learning under Weak Supervision with Performance GuaranteesAlessio Mazzetto, Cyrus Cousins, Dylan Sam, Stephen H. Bach 等ICML 2021 · 被引用 39 次
相关 Paper
- Learning Hyper Label Model for Programmatic Weak SupervisionRenzhi Wu, Shen-En Chen, Jieyu Zhang, Xu ChuICLR 2023 · 被引用 2 次
- Universalizing Weak SupervisionChangho Shin, Winfred Li, Harit Vishwakarma, Nicholas Carl Roberts 等ICLR 2022 · 被引用 35 次
- Learning from weak labelers as constraintsVishwajeet Agrawal, Rattana Pukdee, Maria-Florina Balcan, Pradeep Kumar RavikumarICLR 2025
- A General Framework for Learning from Weak SupervisionHao Chen, Jindong Wang, Lei Feng, Xiang Li 等ICML 2024 · 被引用 13 次
- End-to-End Weak SupervisionSalva Rühling Cachay, Benedikt Boecking, Artur DubrawskiNeurIPS 2021 · 被引用 48 次
