Investigating label suggestions for opinion mining in German Covid-19 social media
Tilman Beck, Ji-Ung Lee, Christina Viehmann, Marcus Maurer, Oliver Quiring, Iryna Gurevych
摘要
This work investigates the use of interactively updated label suggestions to improve upon the efficiency of gathering annotations on the task of opinion mining in German Covid-19 social media data. We develop guidelines to conduct a controlled annotation study with social science students and find that suggestions from a model trained on a small, expert-annotated dataset already lead to a substantial improvement -in terms of inter-annotator agreement (+.14 Fleiss' κ) and annotation quality -compared to students that do not receive any label suggestions. We further find that label suggestions from interactively trained models do not lead to an improvement over suggestions from a static model. Nonetheless, our analysis of suggestion bias shows that annotators remain capable of reflecting upon the suggested label in general. Finally, we confirm the quality of the annotated data in transfer learning experiments between different annotator groups. To facilitate further research in opinion mining on social media data, we release our collected data consisting of 200 expert and 2,785 student annotations. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Hierarchical and Incremental Structural Entropy Minimization for Unsupervised Social Event DetectionYuwei Cao, Hao Peng, Zhengtao Yu, Philip S. YuAAAI 2024 · 被引用 56 次
- Unsupervised Extractive Summarization of Emotion TriggersTiberiu Sosea, Hongli Zhan, Junyi Jessy Li, Cornelia CarageaACL 2023 · 被引用 3 次
- Automatic Identification and Classification of Bragging in Social MediaMali Jin, Daniel Preotiuc-Pietro, A. Seza Dogruöz, Nikolaos AletrasACL 2022
- Structural Entropy Guided Incremental Learning for Open-World Multimodal Social Event DetectionZhiwei Yang, Haimei Qin, Xiaoyan Yu, Hao Peng 等AAAI 2026
它引用的顶会 Paper3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- From Zero to Hero: Human-In-The-Loop Entity Linking in Low Resource DomainsJan-Christoph Klie, Richard Eckart de Castilho, Iryna GurevychACL 2020 · 被引用 42 次
- From Arguments to Key Points: Towards Automatic Argument SummarizationRoy Bar-Haim, Lilach Eden, Roni Friedman, Yoav Kantor 等ACL 2020 · 被引用 5 次
相关 Paper
- Changes in European Solidarity Before and During COVID-19: Evidence from a Large Crowd- and Expert-Annotated Twitter DatasetAlexandra Ils, Dan Liu, Daniela Grunow, Steffen EgerACL 2021
- Stance Detection in COVID-19 TweetsKyle Glandt, Sarthak Khanal, Yingjie Li, Doina Caragea 等ACL 2021
- Noise Correction on Subjective DatasetsUthman Jinadu, Yi DingACL 2024 · 被引用 2 次
- Interactive Weak Supervision: Learning Useful Heuristics for Data LabelingBenedikt Boecking, Willie Neiswanger, Eric P. Xing, Artur DubrawskiICLR 2021 · 被引用 8 次
- Subjective Crowd Disagreements for Subjective Data: Uncovering Meaningful CrowdOpinion with Population-level LearningTharindu Cyril Weerasooriya, Sarah Luger, Saloni Poddar, Ashiqur R. KhudaBukhsh 等ACL 2023 · 被引用 2 次
