Hard vs. Noise: Resolving Hard-Noisy Sample Confusion in Recommender Systems via Large Language Models
Tianrui Song, Wen-Shuo Chao, Hao Liu
Abstract
Implicit feedback, employed in training recommender systems, unavoidably confronts noise due to factors such as misclicks and position bias. Previous studies have attempted to identify noisy samples through their diverged data patterns, such as higher loss values, and mitigate their influence through sample dropping or reweighting. However, we observed that noisy samples and hard samples display similar patterns, leading to hard-noisy confusion issue. Such confusion is problematic as hard samples are vital for modeling user preferences. To solve this problem, we propose LLMHNI framework, leveraging two auxiliary user-item relevance signals generated by Large Language Models (LLMs) to differentiate hard and noisy samples. LLMHNI obtains user-item semantic relevance from LLM-encoded embeddings, which is used in negative sampling to select hard negatives while filtering out noisy false negatives. An objective alignment strategy is proposed to project LLM-encoded embeddings, originally for general language tasks, into a representation space optimized for user-item relevance modeling. LLMHNI also exploits LLM-inferred logical relevance within user-item interactions to identify hard and noisy samples. These LLM-inferred interactions are integrated into the interaction graph and guide denoising with cross-graph contrastive alignment. To eliminate the impact of unreliable interactions induced by LLM hallucination, we propose a graph contrastive learning strategy that aligns representations from randomly edge-dropped views to suppress unreliable edges. Empirical results demonstrate that LLMHNI significantly improves denoising and recommendation performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b32b34ca-3c4f-4aca-a5d3-727d38220b07Builds on17
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Self-supervised Graph Learning for RecommendationJiancan Wu, Xiang Wang, Fuli Feng, Xiangnan He et al.SIGIR 2021 · 1,476 citations
- Are Graph Augmentations Necessary?: Simple Graph Contrastive Learning for RecommendationJunliang Yu, Hongzhi Yin, Xin Xia, Tong Chen et al.SIGIR 2022 · 658 citations
- Representation Learning with Large Language Models for RecommendationXubin Ren, Wei Wei, Lianghao Xia, Lixin Su et al.WWW 2024 · 385 citations
- Clicks can be Cheating: Counterfactual Recommendation for Mitigating Clickbait IssueWenjie Wang, Fuli Feng, Xiangnan He, Hanwang Zhang et al.SIGIR 2021 · 173 citations
Related papers
- Unleashing the Power of Large Language Model for Denoising RecommendationShuyao Wang, Zhi Zheng, Yongduo Sui, Hui XiongWWW 2025 · 18 citations
- CCLRec: Consensus-driven Contrastive Learning for LLM-enhanced Graph RecommendationTing Guo, Dongyu Pei, Litiao Qiu, Xiaoying Liao et al.ICML 2026
- Semantic Enhanced Heterogeneous Hypergraph Network for Collaborative FilteringMingtao Xu, Wei Wei, Peixuan Yang, Hulong WuAAAI 2025 · 4 citations
- Bridging the User-side Knowledge Gap in Knowledge-aware Recommendations with Large Language ModelsZheng Hu, Zhe Li, Ziyun Jiao, Satoshi Nakagawa et al.AAAI 2025 · 17 citations
- Towards S²-Challenges Underlying LLM-Based Augmentation for Personalized News RecommendationShicheng Wang, Hengzhu Tang, Li Gao, Shu Guo et al.AAAI 2025 · 2 citations
