Can Pre-trained Language Models Interpret Similes as Smart as Human?
Qianyu He, Sijie Cheng, Zhixu Li, Rui Xie, Yanghua Xiao
Abstract
Simile interpretation is a crucial task in natural language processing. Nowadays, pre-trained language models (PLMs) have achieved stateof-the-art performance on many tasks. However, it remains under-explored whether PLMs can interpret similes or not. In this paper, we investigate the ability of PLMs in simile interpretation by designing a novel task named Simile Property Probing, i.e., to let the PLMs infer the shared properties of similes. We construct our simile property probing datasets from both general textual corpora and humandesigned questions, containing 1,633 examples covering seven main categories. Our empirical study based on the constructed datasets shows that PLMs can infer similes' shared properties while still underperforming humans. To bridge the gap with human performance, we additionally design a knowledge-enhanced training objective by incorporating the simile knowledge into PLMs via knowledge embedding methods. Our method results in a gain of 8.58% in the probing task and 1.37% in the downstream task of sentiment classification. The datasets and code are publicly available at https://github.com/Abbey4799/PLMs-Interpret-Simile .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 189d88e2-042d-4511-ac2e-6e872f34453fCited by top-tier papers2
- Fantastic Expressions and Where to Find Them: Chinese Simile Generation with Multiple ConstraintsKexin Yang, Dayiheng Liu, Wenqiang Lei, Baosong Yang et al.ACL 2023 · 3 citations
- Metaphor Reasoning is Meta-reasoningQianyu He, Junting Lu, Yikai Zhang, Siyu Yuan et al.ACL 2026
Builds on7
- Evaluating Commonsense in Pre-Trained Language ModelsXuhui Zhou, Yue Zhang, Leyang Cui, Dandan HuangAAAI 2020 · 198 citations
- Coreferential Reasoning Learning for Language RepresentationDeming Ye, Yankai Lin, Jiaju Du, Zhenghao Liu et al.EMNLP 2020 · 164 citations
- Knowledge-Driven Distractor Generation for Cloze-Style Multiple Choice QuestionsSiyu Ren, Kenny Q. ZhuAAAI 2021 · 62 citations
- Generating similes effortlessly like a Pro: A Style Transfer Approach for Simile GenerationTuhin Chakrabarty, Smaranda Muresan, Nanyun PengEMNLP 2020 · 46 citations
- Neural Simile Recognition with Cyclic Multitask Learning and Local AttentionJiali Zeng, Linfeng Song, Jinsong Su, Jun Xie et al.AAAI 2020 · 26 citations
Related papers
- Probing Simile Knowledge from Pre-trained Language ModelsWeijie Chen, Yongzhu Chang, Rongsheng Zhang, Jiashu Pu et al.ACL 2022
- Probing Linguistic Information for Logical Inference in Pre-trained Language ModelsZeming Chen, Qiyue GaoAAAI 2022 · 11 citations
- COPEN: Probing Conceptual Knowledge in Pre-trained Language ModelsHao Peng, Xiaozhi Wang, Shengding Hu, Hailong Jin et al.EMNLP 2022 · 16 citations
- AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated PromptsTaylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace et al.EMNLP 2020 · 1,162 citations
- Metaphors in Pre-Trained Language Models: Probing and Generalization Across Datasets and LanguagesEhsan Aghazadeh, Mohsen Fayyaz, Yadollah YaghoobzadehACL 2022
