Lune

ICDE2026顶会

Improving Data Imputation Through a Tuned Strategy for Dependency Discovery

Bernardo Breve, Loredana Caruccio, Tullio Pizzuti, Giuseppe Polese

2026年份

摘要

Data Imputation approaches aim at improving the quality of data trying to infer values often missing into data. Among others, dependency-based imputation approaches exploit data relationships among attributes, such as Relaxed Functional Dependencies relaxing on the attribute comparison (RFDcs)\left(\text{RFD}_{c} \mathrm{s}\right), to provide semantically coherent and less biased imputations by exploiting similar candidate tuples. However, according to the possibility of considering variable similarity threshold combinations, existing RFD c_{c} discovery algorithms can limit imputation quality or become impractical for big datasets. In this paper, we propose triard, a tuned strategy for dependency discovery that iteratively adjusts similarity thresholds based on imputation performance, thus selecting the most effective RFDcs\mathbf{R F D}_{c} \mathbf{s} for selecting tuple candidates to impute missing values. Experimental results on 20 real-world datasets show the improvement in both imputation performance and execution time with respect to the RFDcs\mathbf{R F D}_{c} \mathbf{s} discovered by the algorithm domino. Moreover, triard resulted in a higher imputation performance, mainly in terms of precision, when compared with other Data Imputation approaches.

问问这篇 Paper

问问你的智能体。

Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。

可以从这些问题问起

智能体调用

Lunesearch_papers

在 Lune 里问

免费开始,无需绑卡

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖