Towards Automatic Discovering of Deep Hybrid Network Architecture for Sequential Recommendation
Mingyue Cheng, Zhiding Liu, Qi Liu, Shenyang Ge, Enhong Chen
摘要
Recent years have witnessed great success in deep learning-based sequential recommendation (SR), which can provide more timely and accurate recommendations. One of the most effective deep SR architectures is to stack high-performance residual blocks, e.g., prevalent self-attentive and convolutional operations, for capturing long- and short-range dependence of sequential behaviors. By carefully revisiting previous models, we observe: 1) simple architecture modification of gating each residual connection can help us train deeper SR models and yield significant improvements; 2) compared with self-attention mechanism, stacking of convolution layers also can cover each item of the whole sequential behaviors and achieve competitive or even superior performance. Guided by these findings, it is meaningful to design a deeper hybrid SR model to ensemble the capacity of both self-attentive and convolutional architectures for SR tasks. In this work, we aim to achieve this goal in the automatic algorithm sense, and propose NASR, an efficient neural architecture search (NAS) method that can automatically select the architecture operation on each layer. Specifically, we firstly design a Table-like search space, involving both self-attentive and convolutional-based SR architectures in a flexible manner. In the search phase, we leverage weight-sharing supernets to encode the entire search space, and further propose to factorize the whole supernet into blocks to ensure the potential candidate SR architectures can be fully trained. Owning to lacking supervisions, we train each block-wise supernet with a self-supervised contrastive optimization scheme, in which the training signals are constructed by conducting data augmentation on original sequential behaviors. The empirical studies show that the discovered deep hybrid network architectures can exhibit substantial improvements over compared baselines, indicating the practicality of searching deep hybrid network architectures on SR tasks. Notably, we show the discovered architecture also enjoys good generalizability and transferability among different datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- FormerTime: Hierarchical Multi-Scale Representations for Multivariate Time Series ClassificationMingyue Cheng, Qi Liu, Zhiding Liu, Zhi Li 等WWW 2023 · 被引用 64 次
- Continuous Input Embedding Size Search For Recommender SystemsYunke Qu, Tong Chen, Xiangyu Zhao, Lizhen Cui 等SIGIR 2023 · 被引用 16 次
- Towards Context-aware Reasoning-enhanced Generative Searching in E-commerceZhiding Liu, Ben Chen, Mingyue Cheng, Enhong Chen 等WWW 2026 · 被引用 3 次
- BLADE: A Behavior-Level Data Augmentation Framework with Dual Fusion Modeling for Multi-Behavior Sequential RecommendationYupeng Li, Mingyue Cheng, Yucong Luo, Yitong Zhou 等AAAI 2026 · 被引用 1 次
它引用的顶会 Paper13
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- ConViT: Improving Vision Transformers with Soft Convolutional Inductive BiasesStéphane d'Ascoli, Hugo Touvron, Matthew L. Leavitt, Ari S. Morcos 等ICML 2021 · 被引用 1,021 次
相关 Paper
- AutoGSR: Neural Architecture Search for Graph-based Session RecommendationJingfan Chen, Guanghui Zhu, Haojun Hou, Chunfeng Yuan 等SIGIR 2022 · 被引用 25 次
- SGAS: Sequential Greedy Architecture SearchGuohao Li, Guocheng Qian, Itzel C. Delgadillo, Matthias Müller 等CVPR 2020
- StackRec: Efficient Training of Very Deep Sequential Recommender Models by Iterative StackingJiachun Wang, Fajie Yuan, Jian Chen, Qingyao Wu 等SIGIR 2021 · 被引用 25 次
- Graph Masked Autoencoder for Sequential RecommendationYaowen Ye, Lianghao Xia, Chao HuangSIGIR 2023 · 被引用 63 次
- NASRec: Weight Sharing Neural Architecture Search for Recommender SystemsTunhou Zhang, Dehua Cheng, Yuchen He, Zhengxing Chen 等WWW 2023 · 被引用 20 次
