Anchor-based Maximum Discrepancy for Relative Similarity Testing
Zhijian Zhou, Liuhua Peng, Xunye Tian, Feng Liu
摘要
The relative similarity testing aims to determine which of the distributions, P or Q, is closer to an anchor distribution U. Existing kernel-based approaches often test the relative similarity with a fixed kernel in a manually specified alternative hypothesis, e.g., Q is closer to U than P. Although kernel selection is known to be important to kernel-based testing methods, the manually specified hypothesis poses a significant challenge for kernel selection in relative similarity testing: Once the hypothesis is specified first, we can always find a kernel such that the hypothesis is rejected. This challenge makes relative similarity testing ill-defined when we want to select a good kernel after the hypothesis is specified. In this paper, we cope with this challenge via learning a proper hypothesis and a kernel simultaneously, instead of learning a kernel after manually specifying the hypothesis. We propose an anchor-based maximum discrepancy (AMD), which defines the relative similarity as the maximum discrepancy between the distances of (U, P) and (U, Q) in a space of deep kernels. Based on AMD, our testing incorporates two phases. In Phase I, we estimate the AMD over the deep kernel space and infer the potential hypothesis. In Phase II, we assess the statistical significance of the potential hypothesis, where we propose a unified testing framework to derive thresholds for tests over different possible hypotheses from Phase I. Lastly, we validate our method theoretically and demonstrate its effectiveness via extensive experiments on benchmark datasets. Codes are publicly available at: https://github.com/zhijianzhouml/AMD.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Are Two Datasets Close Enough With Statistical Significance? A Kernel Distributional Closeness Testing ApproachZhijian Zhou, Liuhua Peng, Xunye Tian, Mingming Gong 等ICML 2026 · 被引用 1 次
- LOTTERY: Learning from Reference-Only Samples in Two-Sample Testing under Size AsymmetryXunye Tian, Zhijian Zhou, Liuhua Peng, Feng LiuICML 2026
它引用的顶会 Paper17
- Data Augmentation Can Improve RobustnessSylvestre-Alvise Rebuffi, Sven Gowal, Dan Andrei Calian, Florian Stimberg 等NeurIPS 2021 · 被引用 427 次
- Learning Deep Kernels for Non-Parametric Two-Sample TestsFeng Liu, Wenkai Xu, Jie Lu, Guangquan Zhang 等ICML 2020 · 被引用 213 次
- Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?Haoang Chi, He Li, Wenjing Yang, Feng Liu 等NeurIPS 2024 · 被引用 124 次
- Maximum Mean Discrepancy Test is Aware of Adversarial AttacksRuize Gao, Feng Liu, Jingfeng Zhang, Bo Han 等ICML 2021 · 被引用 77 次
- MMD-Fuse: Learning and Combining Kernels for Two-Sample Testing Without Data SplittingFelix Biggs, Antonin Schrab, Arthur GrettonNeurIPS 2023 · 被引用 49 次
相关 Paper
- MMD Graph Kernel: Effective Metric Learning for Graphs via Maximum Mean DiscrepancyYan Sun, Jicong FanICLR 2024 · 被引用 17 次
- Learning Kernel Tests Without Data SplittingJonas M. Kübler, Wittawat Jitkrittum, Bernhard Schölkopf, Krikamol MuandetNeurIPS 2020 · 被引用 27 次
- DUAL: Learning Diverse Kernels for Aggregated Two-sample and Independence TestingZhijian Zhou, Xunye Tian, Liuhua Peng, Chao Lei 等NeurIPS 2025 · 被引用 8 次
- Meta Two-Sample Testing: Learning Kernels for Testing with Limited DataFeng Liu, Wenkai Xu, Jie Lu, Danica J. SutherlandNeurIPS 2021 · 被引用 30 次
- Kernel-based Maximum-of-difference Test for Two-sample ComparisonDan Pu, Tianyi Zhu, Yao Yan, Wei LanICML 2026
