DUAL: Learning Diverse Kernels for Aggregated Two-sample and Independence Testing
Zhijian Zhou, Xunye Tian, Liuhua Peng, Chao Lei, Antonin Schrab, Danica J. Sutherland, Feng Liu
摘要
To adapt kernel two-sample and independence testing to complex structured data, aggregation of multiple kernels is frequently employed to boost testing power compared to single-kernel tests. However, we observe a phenomenon that directly maximizing multiple kernel-based statistics may result in highly similar kernels that capture highly overlapping information, limiting the effectiveness of aggregation. To address this, we propose an aggregated statistic that explicitly incorporates kernel diversity based on the covariance between different kernels. Moreover, we identify a fundamental challenge: a trade-off between the diversity among kernels and the test power of individual kernels, i.e., the selected kernels should be both effective and diverse. This motivates a testing framework with selection inference, which leverages information from the training phase to select kernels with strong individual performance from the learned diverse kernel pool. We provide rigorous theoretical statements and proofs to show the consistency on the test power and control of Type-I error, along with asymptotic analysis of the proposed statistics. Lastly, we conducted extensive empirical experiments demonstrating the superior performance of our proposed approach across various benchmarks for both twosample and independence testing. ¶
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Are Two Datasets Close Enough With Statistical Significance? A Kernel Distributional Closeness Testing ApproachZhijian Zhou, Liuhua Peng, Xunye Tian, Mingming Gong 等ICML 2026 · 被引用 1 次
- LOTTERY: Learning from Reference-Only Samples in Two-Sample Testing under Size AsymmetryXunye Tian, Zhijian Zhou, Liuhua Peng, Feng LiuICML 2026
- Adaptive Multiscale Binary Expansion Tests for IndependenceYang Yang, Duo Zheng, Sandeep Jain, Kai Zhang 等ICML 2026
它引用的顶会 Paper19
- Learning Deep Kernels for Non-Parametric Two-Sample TestsFeng Liu, Wenkai Xu, Jie Lu, Guangquan Zhang 等ICML 2020 · 被引用 213 次
- Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?Haoang Chi, He Li, Wenjing Yang, Feng Liu 等NeurIPS 2024 · 被引用 124 次
- Self-Supervised Learning with Kernel Dependence MaximizationYazhe Li, Roman Pogodin, Danica J. Sutherland, Arthur GrettonNeurIPS 2021 · 被引用 107 次
- Maximum Mean Discrepancy Test is Aware of Adversarial AttacksRuize Gao, Feng Liu, Jingfeng Zhang, Bo Han 等ICML 2021 · 被引用 77 次
- Deep Unlearning via Randomized Conditionally Independent HessiansRonak Mehta, Sourav Pal, Vikas Singh, Sathya N. RaviCVPR 2022 · 被引用 50 次
相关 Paper
- Practical Kernel Selection for Kernel-based Conditional Independence TestWenjie Wang, Mingming Gong, Biwei Huang, James Bailey 等NeurIPS 2025 · 被引用 2 次
- Meta Two-Sample Testing: Learning Kernels for Testing with Limited DataFeng Liu, Wenkai Xu, Jie Lu, Danica J. SutherlandNeurIPS 2021 · 被引用 30 次
- MMD-Fuse: Learning and Combining Kernels for Two-Sample Testing Without Data SplittingFelix Biggs, Antonin Schrab, Arthur GrettonNeurIPS 2023 · 被引用 49 次
- Kernel-based Maximum-of-difference Test for Two-sample ComparisonDan Pu, Tianyi Zhu, Yao Yan, Wei LanICML 2026
- Diversity-Aware Recursive Feature Multiple Kernel Learningnan cao, Xu Zhao, Teng ZhangICML 2026
