Unleashing the Potential of Two-Tower Models: Diffusion-Based Cross-Interaction for Large-Scale Matching
Yihan Wang, Fei Xiong, Zhexin Han, Qi Song, Kaiqiao Zhan, Ben Wang
摘要
Two-tower models are widely adopted in the industrial-scale matching stage across a broad range of application domains, such as content recommendations, advertisement systems, and search engines. This model efficiently handles large-scale candidate item screening by separating user and item representations. However, the decoupling network also leads to a neglect of potential information interaction between the user and item representations. Current state-of-the-art (SOTA) approaches include adding a shallow fully connected layer(i.e., COLD), which is limited by performance and can only be used in the ranking stage. For performance considerations, another approach attempts to capture historical positive interaction information from the other tower by regarding them as the input features(i.e., DAT). Later research showed that the gains achieved by this method are still limited because of lacking the guidance on the next user intent. To address the aforementioned challenges, we propose a "cross-interaction decoupling architecture" within our matching paradigm. This user-tower architecture leverages a diffusion module to reconstruct the next positive intention representation and employs a mixed-attention module to facilitate comprehensive cross-interaction. During the next positive intention generation, we further enhance the accuracy of its reconstruction by explicitly extracting the temporal drift within user behavior sequences. Experiments on two real-world datasets and one industrial dataset demonstrate that our method outperforms the SOTA two-tower models significantly, and our diffusion approach outperforms other generative models in reconstructing item representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERTOmar Khattab, Matei ZahariaSIGIR 2020 · 被引用 1,246 次
- Diffusion Recommender ModelWenjie Wang, Yiyan Xu, Fuli Feng, Xinyu Lin 等SIGIR 2023 · 被引用 281 次
相关 Paper
- Long-Sequence Recommendation Models Need Decoupled EmbeddingsNingya Feng, Junwei Pan, Jialong Wu, Baixu Chen 等ICLR 2025
- Domain-Level Disentanglement Framework Based on Information Enhancement for Cross-Domain Cold-Start RecommendationNian Rong, Fei Xiong, Shirui Pan, Guixun Luo 等AAAI 2025 · 被引用 2 次
- De-collapsing User Intent: Adaptive Diffusion Augmentation with Mixture-of-Experts for Sequential RecommendationXiaoxi Cui, Chao Zhao, Yurong Cheng, Xiangmin ZhouAAAI 2026
- A Learnable Fully Interacted Two-Tower Model for Pre-Ranking SystemChao Xiong, Xianwen Yu, Wei Xu, Lei Cheng 等SIGIR 2025 · 被引用 3 次
- Beyond Two-Tower Matching: Learning Sparse Retrievable Cross-Interactions for RecommendationLiangcai Su, Fan Yan, Jieming Zhu, Xi Xiao 等SIGIR 2023 · 被引用 11 次
