Interpreting Open-Domain Modifiers: Decomposition of Wikipedia Categories into Disambiguated Property-Value Pairs
Marius Pasca
Abstract
This paper proposes an open-domain method for automatically annotating modifier constituents ("20th-century") within Wikipedia categories ("20th-century male writers") with properties ("date of birth"). The annotations offer a semantically-anchored understanding of the role of the constituents in defining the underlying meaning of the categories. In experiments over an evaluation set of Wikipedia categories, the proposed method annotates constituent modifiers as semanticallyanchored properties, rather than as mere strings in a previous method. It does so at a better trade-off between precision and recall.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Span Model for Open Information Extraction on Accurate CorpusJunlang Zhan, Hai ZhaoAAAI 2020 · 90 citations
- Open Knowledge Enrichment for Long-tail EntitiesErmei Cao, Difeng Wang, Jiacheng Huang, Wei HuWWW 2020 · 51 citations
- Hypernym Detection Using Strict Partial Order NetworksSarthak Dash, Md. Faisal Mahbub Chowdhury, Alfio Gliozzo, Nandana Mihindukulasooriya et al.AAAI 2020 · 30 citations
Related papers
- Automatically Labeling Low Quality Content on Wikipedia By Leveraging Patterns in Editing BehaviorsSumit Asthana, Sabrina Tobar Thommel, Aaron Lee Halfaker, Nikola BanovicCSCW 2021 · 9 citations
- Geographic Information Retrieval Using Wikipedia ArticlesAmir Krause, Sara CohenWWW 2023 · 6 citations
- Wiki2Prop: A Multimodal Approach for Predicting Wikidata Properties from WikipediaMichael Luggen, Julien Audiffren, Djellel Eddine Difallah, Philippe Cudré-MaurouxWWW 2021 · 16 citations
- WikiSQE: A Large-Scale Dataset for Sentence Quality Estimation in WikipediaKenichiro Ando, Satoshi Sekine, Mamoru KomachiAAAI 2024 · 5 citations
- ToTTo: A Controlled Table-To-Text Generation DatasetAnkur P. Parikh, Xuezhi Wang, Sebastian Gehrmann, Manaal Faruqui et al.EMNLP 2020 · 69 citations
