Scalable Subset Sampling with Neural Conditional Poisson Networks
Adeel Pervez, Phillip Lippe, Efstratios Gavves
Abstract
A number of problems in learning can be formulated in terms of the basic primitive of sampling k elements out of a universe of n elements. This subset sampling operation cannot directly be included in differentiable models, and approximations are essential. Current approaches take an order sampling approach to sampling subsets and depend on differentiable approximations of the Top-k operator for selecting the largest k elements from a set. We present a simple alternative method for sampling subsets based on conditional Poisson sampling. Unlike order sampling approaches, the complexity of the proposed method is independent of the subset size, which makes the method scalable to large subset sizes. We adapt the procedure to make it efficient and amenable to discrete gradient approximations for use in differentiable models. Furthermore, the method allows the subset size parameter k to be differentiable. We validate our approach extensively, on image and text model explanation, image subsampling and stochastic k-nearest neighbor tasks outperforming existing methods in accuracy, efficiency and scalability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- DynaAct: Large Language Model Reasoning with Dynamic Action SpacesXueliang Zhao, Wei Wu, Jian Guan, Qintong Li et al.NeurIPS 2025 · 2 citations
- Less is More: Fewer Interpretable Region via Submodular Subset SelectionRuoyu Chen, Hua Zhang, Siyuan Liang, Jingzhi Li et al.ICLR 2024
- SFESS: Score Function Estimators for k-Subset SamplingKlas Wijk, Ricardo Vinuesa, Hossein AzizpourICLR 2025
Builds on5
- Differentiable Top-k with Optimal TransportYujia Xie, Hanjun Dai, Minshuo Chen, Bo Dai et al.NeurIPS 2020 · 124 citations
- Gradient Estimation with Stochastic Softmax TricksMax B. Paulus, Dami Choi, Daniel Tarlow, Andreas Krause et al.NeurIPS 2020 · 104 citations
- Deep probabilistic subsampling for task-adaptive compressed sensingIris A. M. Huijben, Bastiaan S. Veeling, Ruud J. G. van SlounICLR 2020 · 47 citations
- How do Decisions Emerge across Layers in Neural Models? Interpretation with Differentiable MaskingNicola De Cao, Michael Sejr Schlichtkrull, Wilker Aziz, Ivan TitovEMNLP 2020 · 13 citations
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
Related papers
- SIMPLE: A Gradient Estimator for k-Subset SamplingKareem Ahmed, Zhe Zeng, Mathias Niepert, Guy Van den BroeckICLR 2023 · 2 citations
- LapSum - One Method to Differentiate Them All: Ranking, Sorting and Top-k SelectionLukasz Struski, Michal B. Bednarczyk, Igor T. Podolak, Jacek TaborICML 2025
- Sampling from a k-DPP without looking at all itemsDaniele Calandriello, Michal Derezinski, Michal ValkoNeurIPS 2020 · 30 citations
- Fiber Monte CarloNick Richardson, Deniz Oktay, Yaniv Ovadia, James C. Bowden et al.ICLR 2024
- Set Based Stochastic SubsamplingBruno Andreis, Seanie Lee, Tuan A. Nguyen, Juho Lee et al.ICML 2022
