Iterative Missing Data Imputation with Model Form Adaptation and Non-Missing Feature Supervision
Hao Wang, Zhengnan Li, Zhichao Chen, Xu Chen, Shuting He, Guangyi Liu, Haoxuan Li, Zhouchen Lin
摘要
Iterative imputation is a prevalent method for missing data imputation, where each feature is imputed iteratively by treating it as a target variable estimated from all other features. However, iterative imputation method suffers from two principal limitations: ❶ it imposes a single parametric model form to impute all features , neglecting the potential for optimal models to vary among features, which risks model misspecification; and ❷ it assumes every feature contains missing values , overlooking the potential presence of non-missing features, termed as oracle features , which are informative for imputation. To address these limitations, we propose kernel point imputation (KPI), a bi-level optimization framework for iterative missing data imputation. At the inner level, KPI adaptively learns the optimal model form for each feature within a reproducing kernel Hilbert space, addressing limitation ❶ . At the outer level, KPI utilizes oracle features as supervisory signals to iteratively refine the imputations, addressing limitation ❷ . Experiments demonstrate that KPI outperforms competitive imputation methods. Code is available at https://github.com/FMLYD/kpi.git .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- CSDI: Conditional Score-based Diffusion Models for Probabilistic Time Series ImputationYusuke Tashiro, Jiaming Song, Yang Song, Stefano ErmonNeurIPS 2021 · 被引用 1,245 次
- Missing Data Imputation using Optimal TransportBoris Muzellec, Julie Josse, Claire Boyer, Marco CuturiICML 2020 · 被引用 179 次
- HyperImpute: Generalized Iterative Imputation with Automatic Model SelectionDaniel Jarrett, Bogdan Cebere, Tennison Liu, Alicia Curth 等ICML 2022 · 被引用 129 次
- MIRACLE: Causally-Aware Imputation via Learning Missing Data MechanismsTrent Kyono, Yao Zhang, Alexis Bellot, Mihaela van der SchaarNeurIPS 2021 · 被引用 105 次
- Propensity Matters: Measuring and Enhancing Balancing for RecommendationHaoxuan Li, Yanghao Xiao, Chunyuan Zheng, Peng Wu 等ICML 2023 · 被引用 55 次
相关 Paper
- Think Twice Before Imputation: Optimizing Data Imputation Order for Machine LearningJiaxuan Zhang, Haitao Yuan, Jianing Si, Nan Jiang 等ICDE 2025
- Missing Data Imputation by Reducing Mutual Information with Rectified FlowsJiahao Yu, Qizhen Ying, Leyang Wang, Ziyue Jiang 等NeurIPS 2025 · 被引用 9 次
- Handling Missing Data with Graph Representation LearningJiaxuan You, Xiaobai Ma, Daisy Yi Ding, Mykel J. Kochenderfer 等NeurIPS 2020 · 被引用 274 次
- Probabilistic Missing Value Imputation for Mixed Categorical and Ordered DataYuxuan Zhao, Alex Townsend, Madeleine UdellNeurIPS 2022 · 被引用 3 次
- RefiDiff: Progressive Refinement Diffusion for Efficient Missing Data ImputationMd. Atik Ahamed, Qiang Ye, Qiang ChengAAAI 2026
