Exploring the Escalation of Source Bias in User, Data, and Recommender System Feedback Loop
Yuqi Zhou, Sunhao Dai, Liang Pang, Gang Wang, Zhenhua Dong, Jun Xu, Ji-Rong Wen
摘要
Recommender systems are essential for information access, allowing users to present their content for recommendation. With the rise of large language models (LLMs), AI-generated content (AIGC), primarily in the form of text, has become a central part of the content ecosystem. As AIGC becomes increasingly prevalent, it is important to understand how it affects the performance and dynamics of recommender systems. To this end, we construct an environment that incorporates AIGC to explore its short-term impact. The results from popular sequential recommendation models reveal that AIGC are ranked higher in the recommender system, reflecting the phenomenon of source bias [13,41]. To further explore the long-term impact of AIGC, we introduce a feedback loop with realistic simulators. The results show that the model's preference for AIGC increases as the user clicks on AIGC rises and the model trains on simulated click data. This leads to two issues: In the short term, bias toward AIGC encourages LLM-based content creation, increasing AIGC content, and causing unfair traffic distribution. From a long-term perspective, our experiments also show that when AIGC dominates the content ecosystem after a feedback loop, it can lead to a decline in recommendation performance. To address these issues, we propose a debiasing method based on L1-loss optimization to maintain long-term content ecosystem balance. In a real-world environment with AIGC generated by mainstream LLMs, our method ensures a balance between AIGC and human-generated content in the ecosystem. The code and dataset are available at https://github.com/Yuqi-Zhou/Rec_SourceBias.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- LLM-Generated Fake News Induces Truth Decay in News Ecosystem: A Case Study on Neural News RecommendationBeizhe Hu, Qiang Sheng, Juan Cao, Yang Li 等SIGIR 2025 · 被引用 7 次
- Generative Ghost: Investigating Ranking Bias Hidden in AI-Generated VideosHaowen Gao, Liang Pang, Shicheng Xu, Leigang Qu 等ACM MM 2025 · 被引用 1 次
- Exploring the Evolvement of User Engagement in Online Creative Community under the Surge of Generative AI: A Case Study of DeviantArtQingyu Guo, Yuqi Zhang, Kangyu Yuan, Changyang He 等CSCW 2025
它引用的顶会 Paper9
- Catastrophic Jailbreak of Open-source LLMs via Exploiting GenerationYangsibo Huang, Samyak Gupta, Mengzhou Xia, Kai Li 等ICLR 2024 · 被引用 481 次
- Representation Learning with Large Language Models for RecommendationXubin Ren, Wei Wei, Lianghao Xia, Lixin Su 等WWW 2024 · 被引用 385 次
- Self-Consuming Generative Models Go MADSina Alemohammad, Josue Casco-Rodriguez, Lorenzo Luzi, Ahmed Imtiaz Humayun 等ICLR 2024 · 被引用 279 次
- Search-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive TasksShicheng Xu, Liang Pang, Huawei Shen, Xueqi Cheng 等WWW 2024 · 被引用 104 次
- Neural Retrievers are Biased Towards LLM-Generated ContentSunhao Dai, Yuqi Zhou, Liang Pang, Weihao Liu 等KDD 2024 · 被引用 26 次
相关 Paper
- Spiral of Silence: How is Large Language Model Killing Information Retrieval? - A Case Study on Open Domain Question AnsweringXiaoyang Chen, Ben He, Hongyu Lin, Xianpei Han 等ACL 2024 · 被引用 8 次
- Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated ImagesShicheng Xu, Danyang Hou, Liang Pang, Jingcheng Deng 等SIGIR 2024 · 被引用 18 次
- The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code GenerationXiaoyu Zhang, Juan Zhai, Shiqing Ma, Qingshuang Bao 等ACL 2025 · 被引用 6 次
- Mitigating Source Bias with LLM AlignmentSunhao Dai, Yuqi Zhou, Liang Pang, Zhuoyang Li 等SIGIR 2025 · 被引用 2 次
- Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity DocumentsHaoyu Wang, Sunhao Dai, Haiyuan Zhao, Liang Pang 等ICLR 2025
