HarassGuard: Detecting Harassment Behaviors in Social Virtual Reality with Vision-Language Models
Junhee Lee, Minseok Kim, Hwanjo Heo, Seungwon Woo, Jinwoo Kim
Abstract
Social Virtual Reality (VR) platforms provide immersive social experiences but also expose users to serious risks of online harassment. Existing safety measures are largely reactive, while proactive solutions that detect harassment behavior during an incident often depend on sensitive biometric data, raising privacy concerns. In this paper, we present HarassGuard, a vision-language model (VLM) based system that detects physical harassment in social VR using only visual input. We construct an IRB-approved harassment vision dataset, apply prompt engineering, and fine-tune VLMs to detect harassment behavior by considering contextual information in social VR. Experimental results demonstrate that HarassGuard achieves competitive performance compared to state-of-the-art baselines (i.e., LSTM/CNN, Transformer), reaching an accuracy of up to 88.09% in binary classification and 68.85% in multi-class classification. Notably, HarassGuard matches these baselines while using significantly fewer fine-tuning samples (200 vs. 1,115), offering unique advantages in contextual reasoning and privacy-preserving detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on7
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Disturbing the Peace: Experiencing and Mitigating Emerging Harassment in Social Virtual RealityGuo Freeman, Samaneh Zamanifard, Divine Maloney, Dane AcenaCSCW 2022 · 163 citations
- Enabling Developers, Protecting Users: Investigating Harassment and Safety in VRAbhinaya S. B., Aafaq Sabir, Anupam DasUSENIX Security 2024 · 16 citations
- De-anonymization Attacks on MetaverseYan Meng, Yuxia Zhan, Jiachun Li, Suguo Du et al.INFOCOM 2023 · 13 citations
- Beyond Mute and Block: Adoption and Effectiveness of Safety Tools in Social VR, from Ubiquitous Harassment to Social SculptingMaheshya Weerasinghe, Shaun Alexander Macdonald, Cristina Fiani, Joseph O'Hagan et al.IEEE VR 2025 · 12 citations
Related papers
- HardenVR: Harassment Detection in Social Virtual RealityNa Wang, Jin Zhou, Jie Li, Bo Han et al.IEEE VR 2024 · 9 citations
- Moderating Illicit Online Image Promotion for Unsafe User Generated Content Games Using Large Vision-Language ModelsKeyan Guo, Ayush Utkarsh, Wenbo Ding, Isabelle Ondracek et al.USENIX Security 2024 · 16 citations
- LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and ModelsLukas Helff, Felix Friedrich, Manuel Brack, Kristian Kersting et al.ICML 2025
- Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language ModelsYongshuo Zong, Ondrej Bohdal, Tingyang Yu, Yongxin Yang et al.ICML 2024 · 140 citations
- Towards Policy-Adaptive Image Guardrail: Benchmark and MethodCaiyong Piao, Zhiyuan Yan, Haoming Xu, Yunzhen Zhao et al.CVPR 2026 · 7 citations
