Defending Against Social Engineering Attacks in the Age of LLMs
Lin Ai, Tharindu Kumarage, Amrita Bhattacharjee, Zizhou Liu, Zheng Hui, Michael Davinroy, James Cook, Laura Cassani, Kirill Trapeznikov, Matthias Kirchner, Arslan Basharat, Anthony Hoogs
Abstract
The proliferation of Large Language Models (LLMs) poses challenges in detecting and mitigating digital deception, as these models can emulate human conversational patterns and facilitate chat-based social engineering (CSE) attacks. This study investigates the dual capabilities of LLMs as both facilitators and defenders against CSE threats. We develop a novel dataset, SEConvo, simulating CSE scenarios in academic and recruitment contexts, and designed to examine how LLMs can be exploited in these situations. Our findings reveal that, while off-the-shelf LLMs generate high-quality CSE content, their detection capabilities are suboptimal, leading to increased operational costs for defense. In response, we propose Con-voSentinel, a modular defense pipeline that improves detection at both the message and the conversation levels, offering enhanced adaptability and cost-effectiveness. The retrievalaugmented module in ConvoSentinel identifies malicious intent by comparing messages to a database of similar conversations, enhancing CSE detection at all stages. Our study highlights the need for advanced strategies to leverage LLMs in cybersecurity. Our code and data are available at this GitHub repository.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Searching for Privacy Risks in LLM Agents via SimulationYanzhe Zhang, Diyi YangICLR 2026 · 21 citations
- Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading ComprehensionLin Ai, Zheng Hui, Zizhou Liu, Julia HirschbergEMNLP 2024 · 4 citations
- Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement LearningZheng Hui, Yijiang River Dong, Sanhanat Sivapiromrat, Ehsan Shareghi et al.ACL 2026 · 4 citations
- TBTrackerX: Fantastic Trigger Bots and Where to Find Malicious Campaigns on XMohammad Majid Akhtar, Rahat Masood, Muhammad Ikram, Salil S. KanhereNDSS 2026 · 1 citation
- From Trust to Compromise: Outcome-Verified LLM Phishing Simulation and Real-Time DefenseTulika Tewari, Nalin Asanka Gamagedara Arachchilage, Jagat Sesh Challa, Dhruv KumarACL 2026
Related papers
- DualSentinel: A Lightweight Framework for Detecting Targeted Attacks in Black-box LLM via Dual Entropy Lull PatternXiaoyi Pang, Xuanyi Hao, Pengyu Liu, Qi Luo et al.USENIX Security 2026
- Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting ScamsGilad Gressel, Rahul Pankajakshan, Shir Rozenfeld, Ling Li et al.USENIX Security 2026
- AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use AgentsHaitao Hu, Peng Chen, Yanpeng Zhao, Yuqi ChenCCS 2025
- On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMsHerun Wan, Minnan Luo, Zhixiong Su, Guang Dai et al.ACL 2025 · 5 citations
- Hidden in Plain Sight: Evaluation of the Deception Detection Capabilities of LLMs in Multimodal SettingsMd Messal Monem Miah, Adrita Anika, Xi Shi, Ruihong HuangACL 2025
