Top-Rank-Focused Adaptive Vote Collection for the Evaluation of Domain-Specific Semantic Models
Pierangelo Lombardo, Alessio Boiardi, Luca Colombo, Angelo Schiavone, Nicolò Tamagnone
Abstract
The growth of domain-specific applications of semantic models, boosted by the recent achievements of unsupervised embedding learning algorithms, demands domain-specific evaluation datasets. In many cases, contentbased recommenders being a prime example, these models are required to rank words or texts according to their semantic relatedness to a given concept, with particular focus on top ranks. In this work, we give a threefold contribution to address these requirements: (i) we define a protocol for the construction, based on adaptive pairwise comparisons, of a relatedness-based evaluation dataset tailored on the available resources and optimized to be particularly accurate in top-rank evaluation; (ii) we define appropriate metrics, extensions of well-known ranking correlation coefficients, to evaluate a semantic model via the aforementioned dataset by taking into account the greater significance of top ranks. Finally, (iii) we define a stochastic transitivity model to simulate semantic-driven pairwise comparisons, which confirms the effectiveness of the proposed dataset construction protocol.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Bradley-Terry Rankings for Recommender Systems Across Dataset TaxonomiesEkaterina Grishina, Stepan L. Kuznetsov, Askar Tsyganov, Ilya Ivanov et al.KDD 2026
- Learning Domain Semantics and Cross-Domain Correlations for Paper RecommendationYi Xie, Yuqing Sun, Elisa BertinoSIGIR 2021 · 15 citations
- On the Emergence of Linear Analogies in Word EmbeddingsDaniel J. Korchinski, Dhruva Karkada, Yasaman Bahri, Matthieu WyartNeurIPS 2025 · 10 citations
- Revisiting the Evaluation Protocol of Knowledge Graph Completion Methods for Link PredictionSudhanshu Tiwari, Iti Bansal, Carlos R. RiveroWWW 2021 · 16 citations
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 459 citations
