End-to-End Weak Supervision
Salva Rühling Cachay, Benedikt Boecking, Artur Dubrawski
Abstract
Aggregating multiple sources of weak supervision (WS) can ease the data-labeling bottleneck prevalent in many machine learning applications, by replacing the tedious manual collection of ground truth labels. Current state of the art approaches that do not use any labeled training data, however, require two separate modeling steps: Learning a probabilistic latent variable model based on the WS sources -- making assumptions that rarely hold in practice -- followed by downstream model training. Importantly, the first step of modeling does not consider the performance of the downstream model. To address these caveats we propose an end-to-end approach for directly learning the downstream model by maximizing its agreement with probabilistic labels generated by reparameterizing previous probabilistic posteriors with a neural network. Our results show improved performance over prior work in terms of end model performance on downstream test sets, as well as in terms of improved robustness to dependencies among weak supervision sources.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers17
- Theoretical Analysis of Weak-to-Strong GeneralizationHunter Lang, David A. Sontag, Aravindan VijayaraghavanNeurIPS 2024 · 59 citations
- Nemo: Guiding and Contextualizing Weak Supervision for Interactive Data ProgrammingCheng-Yu Hsieh, Jieyu Zhang, Alexander J. RatnerVLDB 2022 · 17 citations
- Losses over Labels: Weakly Supervised Learning via Direct Loss ConstructionDylan Sam, J. Zico KolterAAAI 2023 · 14 citations
- Characterizing the Impacts of Semi-supervised Learning for Weak SupervisionJeffrey Li, Jieyu Zhang, Ludwig Schmidt, Alexander J. RatnerNeurIPS 2023 · 9 citations
- CARE: Confounder-Aware Aggregation for Reliable LLM EvaluationJitian Zhao, Changho Shin, Tzu-Heng Huang, Satya Sai Srinath Namburi GNVV et al.ICML 2026 · 7 citations
Builds on6
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Understanding self-supervised learning dynamics without contrastive pairsYuandong Tian, Xinlei Chen, Surya GanguliICML 2021 · 338 citations
- Fast and Three-rious: Speeding Up Weak Supervision with Triplet MethodsDaniel Y. Fu, Mayee F. Chen, Frederic Sala, Sarah M. Hooper et al.ICML 2020 · 130 citations
- Learning from Rules Generalizing Labeled ExemplarsAbhijeet Awasthi, Sabyasachi Ghosh, Rasna Goyal, Sunita SarawagiICLR 2020 · 93 citations
- Interactive Weak Supervision: Learning Useful Heuristics for Data LabelingBenedikt Boecking, Willie Neiswanger, Eric P. Xing, Artur DubrawskiICLR 2021 · 8 citations
Related papers
- Generative Modeling Helps Weak Supervision (and Vice Versa)Benedikt Boecking, Nicholas Carl Roberts, Willie Neiswanger, Stefano Ermon et al.ICLR 2023 · 1 citation
- Learning Hyper Label Model for Programmatic Weak SupervisionRenzhi Wu, Shen-En Chen, Jieyu Zhang, Xu ChuICLR 2023 · 2 citations
- Creating Training Sets via Weak Indirect SupervisionJieyu Zhang, Bohan Wang, Xiangchen Song, Yujing Wang et al.ICLR 2022 · 17 citations
- Understanding Programmatic Weak Supervision via Source-aware Influence FunctionJieyu Zhang, Haonan Wang, Cheng-Yu Hsieh, Alexander J. RatnerNeurIPS 2022 · 13 citations
- Amortized Variational Inference for Partial-Label Learning: A Probabilistic Approach to Label DisambiguationTobias Fuchs, Nadja KleinICML 2026
