G2M: A Generalized Gaussian Mirror Method to Boost Feature Selection Power
Hongyu Shen, Zhizhen Jane Zhao
摘要
Recent advances in false discovery rate (FDR)-controlled feature selection methods have improved reliability by effectively limiting false positives, making them wellsuited for complex applications. A popular FDR-controlled framework called data splitting uses the "mirror statistics" to select features. However, we find that the unit variance assumption on mirror statistics could potentially limit the feature selection power. To address this, we generalize the mirror statistics in the Gaussian mirror framework and introduce a new approach called "generalized Gaussian mirror" (G 2 M), which adaptively learns the variance and forms new test statistics. We demonstrate both theoretically and empirically that the proposed test statistics achieve higher power than those of Gaussian mirror and data splitting. Comparisons with other FDR-controlled frameworks on synthetic, semi-synthetic, and real datasets highlight the superior performance of the G 2 M method in achieving higher power while maintaining FDR control. These findings suggest the potential for the G 2 M method for practical applications in real-world problems. Code is available at: https://github.com/skyve2012/G2M.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Deep Direct Likelihood KnockoffsMukund Sudarshan, Wesley Tansey, Rajesh RanganathNeurIPS 2020 · 被引用 25 次
- Normalizing Flows for Knockoff-free Controlled Feature SelectionDerek Hansen, Brian Manzo, Jeffrey RegierNeurIPS 2022 · 被引用 8 次
- A Conditional Randomization Test for Sparse Logistic Regression in High-DimensionBinh T. Nguyen, Bertrand Thirion, Sylvain ArlotNeurIPS 2022 · 被引用 7 次
- DeepDRK: Deep Dependency Regularized Knockoff for Feature SelectionHongyu Shen, Yici Yan, Zhizhen Jane ZhaoNeurIPS 2024 · 被引用 2 次
相关 Paper
- Split Group Knockoffs: Controlling False Discovery Rate in Transformational Group SparsitySiqi Chen, Yachen Gao, Yanwei Fu, Xinwei SunICML 2026
- False Discovery Proportion control for aggregated KnockoffsAlexandre Blain, Bertrand Thirion, Olivier Grisel, Pierre NeuvialNeurIPS 2023 · 被引用 4 次
- Multivariate Conformal SelectionTian Bai, Yue Zhao, Xiang Yu, Archer Y. YangICML 2025
- An Online Statistical Framework for Out-of-Distribution DetectionXinsong Ma, Xin Zou, Weiwei LiuICML 2025
- A unified framework for bandit multiple testingZiyu Xu, Ruodu Wang, Aaditya RamdasNeurIPS 2021 · 被引用 22 次
