ePiC: Employing Proverbs in Context as a Benchmark for Abstract Language Understanding
Sayan Ghosh, Shashank Srivastava
摘要
While large language models have shown exciting progress on several NLP benchmarks, evaluating their ability for complex analogical reasoning remains under-explored. Here, we introduce a high-quality crowdsourced dataset of narratives for employing proverbs in context as a benchmark for abstract language understanding. The dataset provides fine-grained annotation of aligned spans between proverbs and narratives, and contains minimal lexical overlaps between narratives and proverbs, ensuring that models need to go beyond surfacelevel reasoning to succeed. We explore three tasks: (1) proverb recommendation and alignment prediction, (2) narrative generation for a given proverb and topic, and (3) identifying narratives with similar motifs. Our experiments show that neural language models struggle on these tasks compared to humans, and these tasks pose multiple learning challenges. NARRATIVE (N2) NARRATIVE (N1) PROVERB (P)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Help Me Write a Story: Evaluating LLMs' Ability to Generate Writing FeedbackHannah Rashkin, Elizabeth Clark, Fantine Huot, Mirella LapataACL 2025 · 被引用 7 次
- AnaloBench: Benchmarking the Identification of Abstract and Long-context AnalogiesXiao Ye, Andrew Wang, Jacob Choi, Yining Lu 等EMNLP 2024 · 被引用 3 次
- Pun Unintended: LLMs and the Illusion of Humor UnderstandingAlessandro Zangari, Matteo Marcuzzo, Andrea Albarelli, Mohammad Taher Pilehvar 等EMNLP 2025 · 被引用 1 次
- TRoTR: A Framework for Evaluating the Re-contextualization of Text ReuseFrancesco Periti, Pierluigi Cassotti, Stefano Montanelli, Nina Tahmasebi 等EMNLP 2024
- Towards a Greek Proverb Atlas: Computational Spatial Exploration and Attribution of Greek ProverbsJohn Pavlopoulos, Panos Louridas, Panagiotis FilosEMNLP 2024
它引用的顶会 Paper6
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Generating similes effortlessly like a Pro: A Style Transfer Approach for Simile GenerationTuhin Chakrabarty, Smaranda Muresan, Nanyun PengEMNLP 2020 · 被引用 46 次
- Neural Simile Recognition with Cyclic Multitask Learning and Local AttentionJiali Zeng, Linfeng Song, Jinsong Su, Jun Xie 等AAAI 2020 · 被引用 26 次
- Continuity of Topic, Interaction, and Query: Learning to Quote in Online ConversationsLingzhi Wang, Jing Li, Xingshan Zeng, Haisong Zhang 等EMNLP 2020 · 被引用 13 次
相关 Paper
- Can language models learn analogical reasoning? Investigating training objectives and comparisons to human performanceMolly R. Petersen, Lonneke van der PlasEMNLP 2023 · 被引用 3 次
- Metaphor Understanding Challenge Dataset for LLMsXiaoyu Tong, Rochelle Choenni, Martha Lewis, Ekaterina ShutovaACL 2024
- StoryAnalogy: Deriving Story-level Analogies from Large Language Models to Unlock Analogical UnderstandingCheng Jiayang, Lin Qiu, Tsz Ho Chan, Tianqing Fang 等EMNLP 2023 · 被引用 8 次
- IMPLI: Investigating NLI Models' Performance on Figurative LanguageKevin Stowe, Prasetya Ajie Utama, Iryna GurevychACL 2022 · 被引用 52 次
- ExPUNations: Augmenting Puns with Keywords and ExplanationsJiao Sun, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone 等EMNLP 2022 · 被引用 7 次
