HORSE: Hierarchical Representation for Large-Scale Neural Subset Selection
Binghui Xie, Yixuan Wang, Yongqiang Chen, Kaiwen Zhou, Yu Li, Wei Meng, James Cheng
摘要
Subset selection tasks, such as anomaly detection and compound selection in AI-assisted drug discovery, are crucial for a wide range of applications. Learning subset-valued functions with neural networks has achieved great success by incorporating permutation invariance symmetry into the architecture. However, existing neural set architectures often struggle to either capture comprehensive information from the superset or address complex interactions within the input. Additionally, they often fail to perform in scenarios where superset sizes surpass available memory capacity. To address these challenges, we introduce the novel concept of the Identity Property , which requires models to integrate information from the originating set, resulting in the development of neural networks that excel at performing effective subset selection from large supersets. Moreover, we present the Hierarchical Representation of Neural Subset Selection (HORSE), an attention-based method that learns complex interactions and retains information from both the input set and the optimal subset supervision signal. Specifically, HORSE enables the partitioning of the input ground set into manageable chunks that can be processed independently and then aggregated, ensuring consistent outcomes across different partitions. Through extensive experimentation, we demonstrate that HORSE significantly enhances neural subset selection performance by capturing more complex information and surpasses the state-of-the-art methods in handling large-scale inputs by a margin of up to 20%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper17
- Coresets for Data-efficient Training of Machine Learning ModelsBaharan Mirzasoleiman, Jeff A. Bilmes, Jure LeskovecICML 2020 · 被引用 494 次
- Equivariant Subgraph Aggregation NetworksBeatrice Bevilacqua, Fabrizio Frasca, Derek Lim, Balasubramaniam Srinivasan 等ICLR 2022 · 被引用 217 次
- Set Functions for Time SeriesMax Horn, Michael Moor, Christian Bock, Bastian Rieck 等ICML 2020 · 被引用 199 次
- Structure-aware Interactive Graph Neural Networks for the Prediction of Protein-Ligand Binding AffinityShuangli Li, Jingbo Zhou, Tong Xu, Liang Huang 等KDD 2021 · 被引用 184 次
- On Learning Sets of Symmetric ElementsHaggai Maron, Or Litany, Gal Chechik, Ethan FetayaICML 2020 · 被引用 148 次
相关 Paper
- Enhancing Neural Subset Selection: Integrating Background Information into Set RepresentationsBinghui Xie, Yatao Bian, Kaiwen Zhou, Yongqiang Chen 等ICLR 2024 · 被引用 1 次
- Learning Neural Set Functions Under the Optimal Subset OracleZijing Ou, Tingyang Xu, Qinliang Su, Yingzhen Li 等NeurIPS 2022 · 被引用 13 次
- Learning Set Functions with Implicit DifferentiationGözde Özcan, Chengzhi Shi, Stratis IoannidisAAAI 2025
- Efficient Data Subset Selection to Generalize Training Across Models: Transductive and Inductive NetworksEeshaan Jain, Tushar Nandy, Gaurav Aggarwal, Ashish Tendulkar 等NeurIPS 2023 · 被引用 30 次
- Graph Neural Networks with Adaptive ReadoutsDavid Buterez, Jon Paul Janet, Steven J. Kiddle, Dino Oglic 等NeurIPS 2022 · 被引用 81 次
