Learning In-context Learning for Named Entity Recognition
Jiawei Chen, Yaojie Lu, Hongyu Lin, Jie Lou, Wei Jia, Dai Dai, Hua Wu, Boxi Cao, Xianpei Han, Le Sun
摘要
Named entity recognition in real-world applications suffers from the diversity of entity types, the emergence of new entity types, and the lack of high-quality annotations. To address the above problems, this paper proposes an in-context learning-based NER approach, which can effectively inject in-context NER ability into PLMs and recognize entities of novel types on-the-fly using only a few demonstrative instances. Specifically, we model PLMs as a meta-function λ instruction, demonstrations, text .M 1 , and a new entity extractor can be implicitly constructed by applying new instruction and demonstrations to PLMs, i.e., (λ.M)(instruction, demonstrations) → F where F will be a new entity extractor, i.e., F: text → entities. To inject the above in-context NER ability into PLMs, we propose a meta-function pre-training algorithm, which pre-trains PLMs by comparing the (instruction, demonstration)-initialized extractor with a surrogate golden extractor. Experimental results on 4 few-shot NER datasets show that our method can effectively inject in-context NER ability into PLMs and significantly outperforms the PLMs+fine-tuning counterparts. * This work was partially done when Jiawei Chen interned at Baidu. † Corresponding authors. 1 This paper represents functions using lambdacalculus (Barendregt, 1992) , and each function is represented as λx,y,z.M , where x, y, z are variables and M is function definition/abstraction. The function can apply to arguments such as (λ x,y,z .M )(x = A, y = B, z = C) (fully applied) or (λ x,y,z .M )(x = A, y = B) (partially applied). Entities: SARS-CoV-2 is virus. COVID-19 is disease.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- PaDeLLM-NER: Parallel Decoding in Large Language Models for Named Entity RecognitionJinghui Lu, Yanjie Wang, Ziwei Yang, Xuejing Liu 等NeurIPS 2024 · 被引用 22 次
- Guideline Learning for In-Context Information ExtractionChaoxu Pang, Yixuan Cao, Qiang Ding, Ping LuoEMNLP 2023 · 被引用 12 次
- Collaborative Evolution: Multi-Round Learning Between Large and Small Language Models for Emergent Fake News DetectionZiyi Zhou, Xiaoming Zhang, Shenghan Tan, Litian Zhang 等AAAI 2025 · 被引用 9 次
- Unified Low-Resource Sequence Labeling by Sample-Aware Dynamic Sparse FinetuningSarkar Snigdha Sarathi Das, Haoran Zhang, Peng Shi, Wenpeng Yin 等EMNLP 2023 · 被引用 2 次
- PMRC: Prompt-Based Machine Reading Comprehension for Few-Shot Named Entity RecognitionJin Huang, Danfeng Yan, Yuanqiang CaiAAAI 2024 · 被引用 2 次
它引用的顶会 Paper15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Calibrate Before Use: Improving Few-shot Performance of Language ModelsZihao Zhao, Eric Wallace, Shi Feng, Dan Klein 等ICML 2021 · 被引用 1,843 次
- Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order SensitivityYao Lu, Max Bartolo, Alastair Moore, Sebastian Riedel 等ACL 2022 · 被引用 1,494 次
- Transformers Learn In-Context by Gradient DescentJohannes von Oswald, Eyvind Niklasson, Ettore Randazzo, João Sacramento 等ICML 2023 · 被引用 729 次
相关 Paper
- Few-Shot Named Entity Recognition: An Empirical Baseline StudyJiaxin Huang, Chunyuan Li, Krishan Subudhi, Damien Jose 等EMNLP 2021 · 被引用 97 次
- Few-Shot Fine-Grained Entity Typing with Automatic Label Interpretation and Instance GenerationJiaxin Huang, Yu Meng, Jiawei HanKDD 2022 · 被引用 17 次
- Good Examples Make A Faster Learner: Simple Demonstration-based Learning for Low-resource NERDong-Ho Lee, Akshen Kadakia, Kangmin Tan, Mahak Agarwal 等ACL 2022 · 被引用 96 次
- ConsistNER: Towards Instructive NER Demonstrations for LLMs with the Consistency of Ontology and ContextChenxiao Wu, Wenjun Ke, Peng Wang, Zhizhao Luo 等AAAI 2024 · 被引用 15 次
- A Multi-Agent LLM Framework for Multi-Domain Low-Resource In-Context NER via Knowledge Retrieval, Disambiguation and Reflective AnalysisWenxuan Mu, Jinzhong Ning, Di Zhao, Yijia ZhangAAAI 2026 · 被引用 1 次
