Mining the uncertainty patterns of humans and models in the annotation of moral foundations and human values
Neele Falk, Gabriella Lapesa
摘要
The NLP community has converged on considering disagreement in annotation (or human label variation, HLV) as a constitutive feature of subjective tasks. This paper makes a further step by investigating the relationship between HLV and model uncertainty, and the impact of linguistic features of the items on both. We focus on the identification of moral foundations (e.g., care, fairness, loyalty) and human values (e.g., be polite, be honest) in text. We select three standard datasets and proceed into two steps. First, we focus on HLV and analyze the linguistic features (complexity, polarity, pragmatic phenomena, lexical choices) that correlate with HLV. Next, we proceed to uncertainty and its relationship to HLV. We experiment with RoBERTa and Flan-T5 in a number of training setups and evaluation metrics that test the calibration of uncertainty to HLV and its relationship to performance beyond majority vote; next, we analyze the impact of linguistic features on uncertainty. We find that RoBERTa with soft loss is better calibrated to HLV, and we find alignment between calibrated models and humans in the features (textual complexity and polarity) triggering variation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Toward a Perspectivist Turn in Ground Truthing for Predictive ComputingFederico Cabitza, Andrea Campagner, Valerio BasileAAAI 2023 · 被引用 236 次
- Jury Learning: Integrating Dissenting Voices into Machine Learning ModelsMitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel 等CHI 2022 · 被引用 134 次
- Quantifying the Persona Effect in LLM SimulationsTiancheng Hu, Nigel CollierACL 2024 · 被引用 22 次
- What does a Text Classifier Learn about Morality? An Explainable Method for Cross-Domain Comparison of Moral RhetoricEnrico Liscio, Oscar Araque, Lorenzo Gatti, Ionut Constantinescu 等ACL 2023 · 被引用 12 次
- Cross-Cultural Analysis of Human Values, Morals, and Biases in Folk TalesWinston Wu, Lu Wang, Rada MihalceaEMNLP 2023 · 被引用 4 次
相关 Paper
- Through the Lens of Split Vote: Exploring Disagreement, Difficulty and Calibration in Legal Case Outcome ClassificationShanshan Xu, T. Y. S. S. Santosh, Oana Ichim, Barbara Plank 等ACL 2024
- Disentangling Subjectivity and Uncertainty for Hate Speech Annotation and Modeling using GazeÖzge Alaçam, Sanne Hoeken, Andreas Säuberli, Hannes Gröner 等EMNLP 2025
- Flip-Flop Consistency: Unsupervised Training for Robustness to Prompt Perturbations in LLMsParsa Hejabi, Elnaz Rahmati, Alireza Salkhordeh Ziabari, Morteza DehghaniACL 2026 · 被引用 1 次
- Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies Between Model Predictions and Human Responses in VQAJian Lan, Diego Frassinelli, Barbara PlankAAAI 2025 · 被引用 3 次
- D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and EvaluationAida Mostafazadeh Davani, Mark Diaz, Dylan K. Baker, Vinodkumar PrabhakaranEMNLP 2024 · 被引用 3 次
