Semi-knockoffs: a model-agnostic conditional independence testing method with finite-sample guarantees
Angel REYERO LOBO, Thirion Bertrand, Pierre Neuvial
摘要
Conditional independence testing (CIT) is essential for reliable scientific discovery. It prevents spurious findings and enables controlled feature selection. Recent CIT methods have used machine learning (ML) models as surrogates of the underlying distribution. However, model-agnostic approaches require a train-test split, which reduces statistical power. We introduce Semi-knockoffs, a CIT method that can accommodate any pre-trained model, avoids this split, and provides valid p-values and false discovery rate (FDR) control for high-dimensional settings. Unlike methods that rely on the model-X assumption (known input distribution), Semiknockoffs only require conditional expectations for continuous variables. This makes the procedure less restrictive and more practical for machine learning integration. To ensure validity when estimating these expectations, we present two new theoretical results of independent interest: (i) stability for regularized models trained with a null feature and (ii) the double-robustness property.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Understanding Global Feature Contributions With Additive Importance MeasuresIan Covert, Scott M. Lundberg, Su-In LeeNeurIPS 2020 · 被引用 476 次
- Penalizing Gradient Norm for Efficiently Improving Generalization in Deep LearningYang Zhao, Hao Zhang, Xiuyuan HuICML 2022 · 被引用 165 次
- Statistically Valid Variable Importance Assessment through Conditional PermutationsAhmad Chamma, Denis A. Engemann, Bertrand ThirionNeurIPS 2023 · 被引用 23 次
相关 Paper
- Integral-based Knockoffs Inference for Partially Linear ModelsHao Wang, Biqin Song, Rushi Lan, Hong ChenAAAI 2026
- Error-Based Knockoffs Inference for Controlled Feature SelectionXuebin Zhao, Hong Chen, Yingjie Wang, Weifu Li 等AAAI 2022 · 被引用 2 次
- Normalizing Flows for Knockoff-free Controlled Feature SelectionDerek Hansen, Brian Manzo, Jeffrey RegierNeurIPS 2022 · 被引用 8 次
- DeepDRK: Deep Dependency Regularized Knockoff for Feature SelectionHongyu Shen, Yici Yan, Zhizhen Jane ZhaoNeurIPS 2024 · 被引用 2 次
- False Discovery Proportion control for aggregated KnockoffsAlexandre Blain, Bertrand Thirion, Olivier Grisel, Pierre NeuvialNeurIPS 2023 · 被引用 4 次
