ToMMeR - Efficient Entity Mention Detection from Large Language Models
Victor Morand, Nadi Tomeh, Josiane Mothe, Benjamin Piwowarski
摘要
Identifying which text spans refer to entities - mention detection - is both foundational for information extraction and a known performance bottleneck. We introduce ToMMeR, a lightweight model (<300K parameters) probing mention detection capabilities from early LLM layers. Across 13 NER benchmarks, ToMMeR achieves 93% recall zero-shot, with an estimated 90% precision under a human-calibrated LLM-judge protocol, showing that ToMMeR rarely produces spurious predictions despite high recall. Cross-model analysis reveals that diverse architectures (14M-15B parameters) converge on similar mention boundaries (DICE>75%), confirming that mention detection emerges naturally from language modeling. When extended with span classification heads, ToMMeR achieves competitive NER performance (80-87% F1 on standard benchmarks). Our work provides evidence that structured entity representations exist in early transformer layers and can be efficiently recovered with minimal parameters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- Locating and Editing Factual Associations in GPTKevin Meng, David Bau, Alex Andonian, Yonatan BelinkovNeurIPS 2022 · 被引用 3,415 次
- CrossNER: Evaluating Cross-Domain Named Entity RecognitionZihan Liu, Yan Xu, Tiezheng Yu, Wenliang Dai 等AAAI 2021 · 被引用 201 次
- Universal Information Extraction as Unified Semantic MatchingJie Lou, Yaojie Lu, Dai Dai, Wei Jia 等AAAI 2023 · 被引用 96 次
- How do Language Models Bind Entities in Context?Jiahai Feng, Jacob SteinhardtICLR 2024 · 被引用 81 次
相关 Paper
- Multimodal Language Models See Better When They Look ShallowerHaoran Chen, Junyan Lin, Xinghao Chen, Yue Fan 等EMNLP 2025
- Seq2seq is All You Need for Coreference ResolutionWenzheng Zhang, Sam Wiseman, Karl StratosEMNLP 2023 · 被引用 5 次
- Mention Memory: incorporating textual knowledge into Transformers through entity mention attentionMichiel de Jong, Yury Zemlyanskiy, Nicholas FitzGerald, Fei Sha 等ICLR 2022 · 被引用 55 次
- ExtEnD: Extractive Entity DisambiguationEdoardo Barba, Luigi Procopio, Roberto NavigliACL 2022
- NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated DataSergei Bogdanov, Alexandre Constantin, Timothée Bernard, Benoît Crabbé 等EMNLP 2024 · 被引用 29 次
