Evaluating Taxonomy Free Character Role Labeling (TF-CRL) in News Stories using Large Language Models
David G. Hobson, Derek Ruths, Andrew Piper
摘要
We introduce Taxonomy-Free Character Role Labeling (TF-CRL); a novel task that assigns open-ended narrative role labels to characters in news stories based on their functional role in the narrative. Unlike fixed taxonomies, TF-CRL enables more nuanced and comparative analysis by generating compositional labels (e.g., Resilient Leader, Scapegoated Visionary). We evaluate several large language models (LLMs) on this task using human preference rankings and ratings across four criteria: faithfulness, relevance, informativeness, and generalizability. LLMs almost uniformly outperform human annotators across all dimensions. We further show how TF-CRL supports rich narrative analysis by revealing novel latent taxonomies and enabling cross-domain narrative comparisons. Our approach offers new tools for studying media portrayals, character framing, and the socio-political impacts of narrative roles at-scale. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP ModelsShangbin Feng, Chan Young Park, Yuhan Liu, Yulia TsvetkovACL 2023 · 被引用 117 次
- Concept Induction: Analyzing Unstructured Text with High-Level Concepts Using LLooMMichelle S. Lam, Janice Teoh, James A. Landay, Jeffrey Heer 等CHI 2024 · 被引用 46 次
- ClusterLLM: Large Language Models as a Guide for Text ClusteringYuwei Zhang, Zihan Wang, Jingbo ShangEMNLP 2023 · 被引用 43 次
- APPLS: Evaluating Evaluation Metrics for Plain Language SummarizationYue Guo, Tal August, Gondy Leroy, Trevor Cohen 等EMNLP 2024 · 被引用 7 次
- Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional WorksXinfeng Yuan, Siyu Yuan, Yuhan Cui, Tianhe Lin 等EMNLP 2024 · 被引用 2 次
相关 Paper
- Who Plays Which Role When? Communication Role Dynamics for Peer Recognition and Team Performance PredictionYifan Song, Wenxuan Wendy Shi, Brian P. Bailey, Tal AugustACL 2026
- My side, your side and the evidence: Discovering aligned actor groups and the narratives they weavePavan Holur, David Chong, Timothy R. Tangherlini, Vwani RoychowdhuryACL 2023 · 被引用 1 次
- Themis: A Reference-free NLG Evaluation Language Model with Flexibility and InterpretabilityXinyu Hu, Li Lin, Mingqi Gao, Xunjian Yin 等EMNLP 2024 · 被引用 2 次
- Story Morals: Surfacing value-driven narrative schemas using large language modelsDavid G. Hobson, Haiqi Zhou, Derek Ruths, Andrew PiperEMNLP 2024 · 被引用 4 次
- Can Large Language Models Be an Alternative to Human Evaluations?David Cheng-Han Chiang, Hung-yi LeeACL 2023 · 被引用 254 次
