Robust Fake News Detection using Large Language Models under Adversarial Sentiment Attacks
Sahar Tahmasebi, Eric Müller-Budack, Ralph Ewerth
Abstract
Misinformation and fake news have become a pressing societal challenge, driving the need for reliable automated detection methods. Prior research has highlighted sentiment as an important signal in fake news detection, either by analyzing which sentiments are associated with fake news or by using sentiment and emotion features for classification. However, this poses a vulnerability since adversaries can manipulate sentiment to evade detectors especially with the advent of large language models (LLMs). A few studies have explored adversarial samples generated by LLMs, but they mainly focus on stylistic features such as writing style of news publishers. Thus, the crucial vulnerability of sentiment manipulation remains largely unexplored. In this paper, we investigate the robustness of state-of-the-art fake news detectors under sentiment manipulation. We introduce AdSent, a sentiment-robust detection framework designed to ensure consistent veracity predictions across both original and sentiment-altered news articles. Specifically, we (1) propose controlled sentiment-based adversarial attacks using LLMs, (2) analyze the impact of sentiment shifts on detection performance. We show that changing the sentiment heavily impacts the performance of fake news detection models, indicating biases towards neutral articles being real, while non-neutral articles are often classified as fake content. (3) We introduce a novel sentiment-agnostic training strategy that enhances robustness against such perturbations. Extensive experiments on three benchmark datasets demonstrate that AdSent significantly outperforms competitive baselines in both accuracy and robustness, while also generalizing effectively to unseen datasets and adversarial scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95e5bdf4-d31b-4e72-8138-e06c5147d52eCited by top-tier papers1
Ask how each one uses itBuilds on11
- Mining Dual Emotion for Fake News DetectionXueyao Zhang, Juan Cao, Xirong Li, Qiang Sheng et al.WWW 2021 · 332 citations
- Can LLM-Generated Misinformation Be Detected?Canyu Chen, Kai ShuICLR 2024 · 270 citations
- Zoom Out and Observe: News Environment Perception for Fake News DetectionQiang Sheng, Juan Cao, Xueyao Zhang, Rundong Li et al.ACL 2022 · 103 citations
- The Surprising Performance of Simple Baselines for Misinformation DetectionKellin Pelrine, Jacob Danovitch, Reihaneh RabbanyWWW 2021 · 79 citations
- Fake News in Sheep's Clothing: Robust Fake News Detection Against LLM-Empowered Style AttacksJiaying Wu, Jiafeng Guo, Bryan HooiKDD 2024 · 69 citations
Related papers
- Adversarial Style Augmentation via Large Language Model for Robust Fake News DetectionSungwon Park, Sungwon Han, Xing Xie, Jae-Gil Lee et al.WWW 2025 · 9 citations
- Model-Agnostic Sentiment Distribution Stability Analysis for Robust LLM-Generated Texts DetectionSiyuan Li, Xi Lin, Guangyan Li, Zehao Liu et al.AAAI 2026
- FACTGUARD: Event-Centric and Commonsense-Guided Fake News DetectionJing He, Han Zhang, Yuanhui Xiao, Wei Guo et al.AAAI 2026
- Mitigating Adversarial Attacks by Transferring LLM-generated Narrative Reasoning for Robust Fake News DetectionMengyang Chen, Lingwei Wei, Wei Zhou, Songlin HuSIGIR 2026
- PHPFND: Detecting Fake News via Post-Hoc Processing of LLMs HallucinationJinke Ma, Jiachen Ma, Wei Zhang, Yong LiuAAAI 2026
