LOTTERY: Learning from Reference-Only Samples in Two-Sample Testing under Size Asymmetry
Xunye Tian, Zhijian Zhou, Liuhua Peng, Feng Liu
摘要
Data-adaptive two-sample testing assesses if two samples come from the same distribution, using a discrepancy learned from the data (e.g., via kernelbased feature representations). Such methods typically rely on data splitting to decouple learning from testing and control type I error. However, this paradigm is ill-suited to few-shot settings with severe sample-size imbalance: abundant reference samples are available, while only a handful of query samples arrive. In this paper, we show how this imbalance can be leveraged constructively. Using abundant reference data, we learn reference-dependent representations that summarize salient structure of the reference distribution and provide informative signals for detecting departures. We incorporate a collection of representation families that capture both global and local structure, and adaptively weight them using only reference samples via an uncertaintyguided principle. Theoretically, we establish permutation-based type I error control and show consistency of the aggregated test: as the sample sizes grow, the test power converges to one whenever the representation set contains at least one consistent representation. Empirically, our aggregation achieves strong performance across a range of benchmarks while retaining type I error control. The code of our LOTTERY is available at github.com/yeager20001118/ lottery-two-sample-testing-paper
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- SSD: A Unified Framework for Self-Supervised Outlier DetectionVikash Sehwag, Mung Chiang, Prateek MittalICLR 2021 · 被引用 410 次
- Explainable Deep One-Class ClassificationPhilipp Liznerski, Lukas Ruff, Robert A. Vandermeulen, Billy Joe Franks 等ICLR 2021 · 被引用 240 次
- Learning Deep Kernels for Non-Parametric Two-Sample TestsFeng Liu, Wenkai Xu, Jie Lu, Guangquan Zhang 等ICML 2020 · 被引用 213 次
- DROCC: Deep Robust One-Class ClassificationSachin Goyal, Aditi Raghunathan, Moksh Jain, Harsha Vardhan Simhadri 等ICML 2020 · 被引用 202 次
- Deep One-Class Classification via Interpolated Gaussian DescriptorYuanhong Chen, Yu Tian, Guansong Pang, Gustavo CarneiroAAAI 2022 · 被引用 139 次
相关 Paper
- DUAL: Learning Diverse Kernels for Aggregated Two-sample and Independence TestingZhijian Zhou, Xunye Tian, Liuhua Peng, Chao Lei 等NeurIPS 2025 · 被引用 8 次
- Realistic evaluation of transductive few-shot learningOlivier Veilleux, Malik Boudiaf, Pablo Piantanida, Ismail Ben AyedNeurIPS 2021 · 被引用 55 次
- Universal Representation Learning from Multiple Domains for Few-shot ClassificationWei-Hong Li, Xialei Liu, Hakan BilenICCV 2021 · 被引用 114 次
- MMD-Fuse: Learning and Combining Kernels for Two-Sample Testing Without Data SplittingFelix Biggs, Antonin Schrab, Arthur GrettonNeurIPS 2023 · 被引用 49 次
- FedFSL-CFRD: Personalized Federated Few-Shot Learning with Collaborative Feature Representation DisentanglementShanfeng Wang, Jianzhao Li, Zaitian Liu, Yourun Zhang 等AAAI 2025 · 被引用 1 次
