CORE: A Few-Shot Company Relation Classification Dataset for Robust Domain Adaptation
Philipp Borchert, Jochen De Weerdt, Kristof Coussement, Arno De Caigny, Marie-Francine Moens
Abstract
We introduce CORE, a dataset for few-shot relation classification (RC) focused on company relations and business entities. CORE includes 4,708 instances of 12 relation types with corresponding textual evidence extracted from company Wikipedia pages. Company names and business entities pose a challenge for few-shot RC models due to the rich and diverse information associated with them. For example, a company name may represent the legal entity, products, people, or business divisions depending on the context. Therefore, deriving the relation type between entities is highly dependent on textual context. To evaluate the performance of state-of-the-art RC models on the CORE dataset, we conduct experiments in the few-shot domain adaptation setting. Our results reveal substantial performance gaps, confirming that models trained on different domains struggle to adapt to CORE. Interestingly, we find that models trained on CORE showcase improved out-of-domain performance, which highlights the importance of high-quality data for robust domain adaptation. Specifically, the information richness embedded in business entities allows models to focus on contextual nuances, reducing their reliance on superficial clues such as relation-specific verbs. In addition to the dataset, we provide relevant code snippets to facilitate reproducibility and encourage further research in the field. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1f5206d3-e083-46ca-89b7-77ca2ee53cc4Cited by top-tier papers1
Ask how each one uses itBuilds on6
- Adversarial Soft Prompt Tuning for Cross-Domain Sentiment AnalysisHui Wu, Xiaodong ShiACL 2022 · 98 citations
- Exploring Task Difficulty for Few-Shot Relation ExtractionJiale Han, Bo Cheng, Wei LuEMNLP 2021 · 74 citations
- Knowledge-Enhanced Domain Adaptation in Few-Shot Relation ClassificationJiawen Zhang, Jiaqi Zhu, Yi Yang, Wandong Shi et al.KDD 2021 · 17 citations
- CiteSum: Citation Text-guided Scientific Extreme Summarization and Domain Adaptation with Limited SupervisionYuning Mao, Ming Zhong, Jiawei HanEMNLP 2022 · 11 citations
- TACRED Revisited: A Thorough Evaluation of the TACRED Relation Extraction TaskChristoph Alt, Aleksandra Gabryszak, Leonhard HennigACL 2020 · 9 citations
Related papers
- Few-NERD: A Few-shot Named Entity Recognition DatasetNing Ding, Guangwei Xu, Yulin Chen, Xiaobin Wang et al.ACL 2021
- Universal Representation Learning from Multiple Domains for Few-shot ClassificationWei-Hong Li, Xialei Liu, Hakan BilenICCV 2021 · 114 citations
- Robust Few-Shot Named Entity Recognition with Boundary Discrimination and Correlation PurificationXiaojun Xue, Chunxia Zhang, Tianxiang Xu, Zhendong NiuAAAI 2024 · 7 citations
- A Multi-Mode Modulator for Multi-Domain Few-Shot ClassificationYanbin Liu, Juho Lee, Linchao Zhu, Ling Chen et al.ICCV 2021 · 43 citations
- Better Few-Shot Relation Extraction with Label Prompt DropoutPeiyuan Zhang, Wei LuEMNLP 2022 · 21 citations
