Paired Examples as Indirect Supervision in Latent Decision Models
Nitish Gupta, Sameer Singh, Matt Gardner, Dan Roth
摘要
Compositional, structured models are appealing because they explicitly decompose problems and provide interpretable intermediate outputs that give confidence that the model is not simply latching onto data artifacts. Learning these models is challenging, however, because end-task supervision only provides a weak indirect signal on what values the latent decisions should take. This often results in the model failing to learn to perform the intermediate tasks correctly. In this work, we introduce a way to leverage paired examples that provide stronger cues for learning latent decisions. When two related training examples share internal substructure, we add an additional training objective to encourage consistency between their latent decisions. Such an objective does not require external supervision for the values of the latent output, or even the end task, yet provides an additional training signal to that provided by individual training examples themselves. We apply our method to improve compositional question answering using neural module networks on the DROP dataset. We explore three ways to acquire paired questions in DROP: (a) discovering naturally occurring paired examples within the dataset, (b) constructing paired examples using templates, and (c) generating paired examples using a question generation model. We empirically demonstrate that our proposed approach improves both in-and outof-distribution generalization and leads to correct latent decision predictions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- On Learning Latent Models with Multi-Instance Weak SupervisionKaifu Wang, Efthymia Tsamoura, Dan RothNeurIPS 2023 · 被引用 19 次
- CREST: A Joint Framework for Rationalization and Counterfactual Text GenerationMarcos V. Treviso, Alexis Ross, Nuno Miguel Guerreiro, André F. T. MartinsACL 2023 · 被引用 7 次
- Imbalances in Neurosymbolic Learning: Characterization and Mitigating StrategiesEfthymia Tsamoura, Kaifu Wang, Dan RothNeurIPS 2025 · 被引用 2 次
- Learning with Instance Bundles for Reading ComprehensionDheeru Dua, Pradeep Dasigi, Sameer Singh, Matt GardnerEMNLP 2021 · 被引用 1 次
- On the Compositional Generalization in Versatile Open-domain DialogueTingchen Fu, Xueliang Zhao, Lemao Liu, Rui YanACL 2023
它引用的顶会 Paper9
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong 等NeurIPS 2020 · 被引用 2,774 次
- Asking and Answering Questions to Evaluate the Factual Consistency of SummariesAlex Wang, Kyunghyun Cho, Mike LewisACL 2020 · 被引用 317 次
- Adversarial Filters of Dataset BiasesRonan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula, Rowan Zellers 等ICML 2020 · 被引用 242 次
- Neural Module Networks for Reasoning over TextNitish Gupta, Kevin Lin, Dan Roth, Sameer Singh 等ICLR 2020 · 被引用 134 次
相关 Paper
- Obtaining Faithful Interpretations from Compositional Neural NetworksSanjay Subramanian, Ben Bogin, Nitish Gupta, Tomer Wolfson 等ACL 2020 · 被引用 5 次
- Detection-Based Intermediate Supervision for Visual Question AnsweringYuhang Liu, Daowan Peng, Wei Wei, Yuanyuan Fu 等AAAI 2024 · 被引用 3 次
- Successive Prompting for Decomposing Complex QuestionsDheeru Dua, Shivanshu Gupta, Sameer Singh, Matt GardnerEMNLP 2022 · 被引用 38 次
- Weakly Supervised Neuro-Symbolic Module Networks for Numerical Reasoning over TextAmrita Saha, Shafiq R. Joty, Steven C. H. HoiAAAI 2022 · 被引用 20 次
- Learning by Analogy: A Causal Framework for Compositional GeneralizationLingjing Kong, Shaoan Xie, Yang Jiao, Yetian Chen 等CVPR 2026
