Polyhedron Attention Module: Learning Adaptive-order Interactions
Tan Zhu, Fei Dou, Xinyu Wang, Jin Lu, Jinbo Bi
摘要
Learning feature interactions can be the key for multivariate predictive modeling. ReLU-activated neural networks create piecewise linear prediction models. Other nonlinear activation functions lead to models with only high-order feature interactions, thus lacking of interpretability. Recent methods incorporate candidate polynomial terms of fixed orders into deep learning, which is subject to the issue of combinatorial explosion, or learn the orders that are difficult to adapt to different regions of the feature space. We propose a Polyhedron Attention Module (PAM) to create piecewise polynomial models where the input space is split into poly-hedrons which define the different pieces and on each piece the hyperplanes that define the polyhedron boundary multiply to form the interactive terms, resulting in interactions of adaptive order to each piece. PAM is interpretable to identify important interactions in predicting a target. Theoretic analysis shows that PAM has stronger expression capability than ReLU-activated networks. Extensive experimental results demonstrate the superior classification performance of PAM on massive datasets of the click-through rate prediction and PAM can learn meaningful interaction effects in a medical problem.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- DCN V2: Improved Deep & Cross Network and Practical Lessons for Web-scale Learning to Rank SystemsRuoxi Wang, Rakesh Shivanna, Derek Zhiyuan Cheng, Sagar Jain 等WWW 2021 · 被引用 793 次
- Adaptive Factorization Network: Learning Adaptive-Order Feature InteractionsWeiyu Cheng, Yanyan Shen, Linpeng HuangAAAI 2020 · 被引用 202 次
- FinalMLP: An Enhanced Two-Stream MLP Model for CTR PredictionKelong Mao, Jieming Zhu, Liangcai Su, Guohao Cai 等AAAI 2023 · 被引用 142 次
- FM2: Field-matrixed Factorization Machines for Recommender SystemsYang Sun, Junwei Pan, Alex Zhang, Aaron FloresWWW 2021 · 被引用 98 次
- Feature Interaction Interpretability: A Case for Explaining Ad-Recommendation Systems via Neural Interaction DetectionMichael Tsang, Dehua Cheng, Hanpeng Liu, Xue Feng 等ICLR 2020 · 被引用 71 次
相关 Paper
- Learning Prescriptive ReLU NetworksWei Sun, Asterios TsiourvasICML 2023 · 被引用 3 次
- TropEx: An Algorithm for Extracting Linear Terms in Deep Neural NetworksMartin Trimmel, Henning Petzka, Cristian SminchisescuICLR 2021 · 被引用 15 次
- EulerNet: Adaptive Feature Interaction Learning via Euler's Formula for CTR PredictionZhen Tian, Ting Bai, Wayne Xin Zhao, Ji-Rong Wen 等SIGIR 2023 · 被引用 58 次
- Scalable Interpretability via PolynomialsAbhimanyu Dubey, Filip Radenovic, Dhruv MahajanNeurIPS 2022 · 被引用 42 次
- HIEN: Hierarchical Intention Embedding Network for Click-Through Rate PredictionZuowu Zheng, Changwang Zhang, Xiaofeng Gao, Guihai ChenSIGIR 2022 · 被引用 16 次
