On the Mechanisms of Collaborative Learning in VAE Recommenders
Long Tung Vuong, Julien Monteil, Hien Dang, Volodymyr Vaskovych, Trung Le, Vu Nguyen
Abstract
Variational Autoencoders (VAEs) are a powerful alternative to matrix factorization for recommendation. A common technique in VAE-based collaborative filtering (CF) consists in applying binary input masking to user interaction vectors, which improves performance but remains underexplored theoretically. In this work, we analyze how collaboration arises in VAE-based CF and show it is governed by latent proximity: we derive a latent sharing radius that informs when an SGD update on one user strictly reduces the loss on another user, with influence decaying as the latent Wasserstein distance increases. We further study the induced geometry: with clean inputs, VAE‑based CF primarily exploits local collaboration between input‑similar users and under‑utilizes global collaboration between far‑but‑related users. We compare two mechanisms that encourage global mixing and characterize their trade‑offs: 172 ‑KL regularization directly tightens the information bottleneck, promoting posterior overlap but risking representational collapse if too large; 173 input masking induces stochastic geometric contractions and expansions, which can bring distant users onto the same latent neighborhood but also introduce neighborhood drift. To preserve user identity while enabling global consistency, we propose an anchor regularizer that aligns user posteriors with item embeddings, stabilizing users under masking and facilitating signal sharing across related items. Our analyses are validated on the Netflix, MovieLens-20M, and Million Song datasets. We also successfully deployed our proposed algorithm on an Amazon streaming platform following a successful online experiment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b8af0444-5dab-42d4-bbab-a1de1dc65d27Builds on5
- Autoencoders that don't overfit towards the IdentityHarald SteckNeurIPS 2020 · 72 citations
- Alleviating Cold-start Problem in CTR Prediction with A Variational Embedding Learning FrameworkXiaoxiao Xu, Chen Yang, Qian Yu, Zhiwei Fang et al.WWW 2022 · 41 citations
- TopicVAE: Topic-aware Disentanglement Representation Learning for Enhanced RecommendationZhiqiang Guo, Guohui Li, Jianjun Li, Huaicong ChenACM MM 2022 · 19 citations
- Collaborative Retrieval for Large Language Model-based Conversational Recommender SystemsYaochen Zhu, Chao Wan, Harald Steck, Dawen Liang et al.WWW 2025 · 15 citations
- Sinkhorn Collaborative FilteringXiucheng Li, Jin Yao Chin, Yile Chen, Gao CongWWW 2021 · 8 citations
Related papers
- Mutually-Regularized Dual Collaborative Variational Auto-encoder for Recommendation SystemsYaochen Zhu, Zhenzhong ChenWWW 2022 · 32 citations
- DR-VAE: Debiased and Representation-enhanced Variational Autoencoder for Collaborative RecommendationFan Wang, Chaochao Chen, Weiming Liu, Minye Lei et al.AAAI 2025 · 8 citations
- Stochastic-Expert Variational Autoencoder for Collaborative FilteringYoon-Sik Cho, Min-hwan OhWWW 2022 · 15 citations
- Consistency Regularization for Variational Auto-EncodersSamarth Sinha, Adji Bousso DiengNeurIPS 2021 · 83 citations
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 59 citations
