VPD-100K: Towards Generalizable and Fine-grained Visual Privacy Protection
Xiaobin Hu, Enpu zuo, Lanping Hu, Kaiwen Yang, Dianshu Liao, Tianyi Zhang, Bo Yin, Yinsi Zhou, Shidong Pan, xiaoyu sun
摘要
Privacy protection has become a critical requirement in the era of ubiquitous visual data sharing, imposing higher demands on efficient and robust privacy detection algorithms. However, current robust detection models are severely hindered by the lack of comprehensive datasets. Existing privacy-oriented datasets often suffer from limited scale, coarse-grained annotations, and narrow domain coverage, failing to capture the intricate details of sensitive information in real-world environments. To bridge this gap, we present a large-scale, fine-grained Visual Privacy Dataset (VPD-100K), designed to facilitate generalized privacy detection. We establish a holistic taxonomy comprising four primary domains: Human Presence, On-Screen Personally Identifiable Information (PII), Physical Identifiers, and Location Indicators, containing 100,000 images annotated with 33 fine-grained classes and over 190,000 object instances. Statistical analysis reveals that our dataset features long-tailed distributions, small object scales, and high visual complexity. These characteristics make the dataset particularly valuable for demanding, unconstrained applications such as live streaming, where actors frequently face unintentional, real-time information leakage. Furthermore, we design an effective frequency-enhance lightweight module consisting of frequency-domain attention fusion and adaptive spectral gating mechanism that breaks the limitations of spatial pixel intensity to better capture the subtle details of sensitive information. Extensive experiments conducted on both diverse image and streaming videos benchmarks consistently demonstrate the effectiveness of our VPD-100K dataset and the well-curated frequency mechanism.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Gold-YOLO: Efficient Object Detector via Gather-and-Distribute MechanismChengcheng Wang, Wei He, Ying Nie, Jianyuan Guo 等NeurIPS 2023 · 被引用 732 次
- FBRT-YOLO: Faster and Better for Real-Time Aerial Image DetectionYao Xiao, Tingfa Xu, Yu Xin, Jianan LiAAAI 2025 · 被引用 115 次
- Disability-First Design and Creation of A Dataset Showing Private Visual Information Collected With People Who Are BlindTanusree Sharma, Abigale Stangl, Lotus Zhang, Yu-Yun Tseng 等CHI 2023 · 被引用 23 次
- Do Streamers Care about Bystanders' Privacy? An Examination of Live Streamers' Considerations and Strategies for Bystanders' Privacy ManagementYanlai Wu, Xinning Gui, Pamela J. Wisniewski, Yao LiCSCW 2023 · 被引用 15 次
- DIPA2: An Image Dataset with Cross-cultural Privacy Perception AnnotationsAnran Xu, Zhongyi Zhou, Kakeru Miyazaki, Ryo Yoshikawa 等UbiComp 2024 · 被引用 15 次
相关 Paper
- Large-scale Video Panoptic Segmentation in the Wild: A BenchmarkJiaxu Miao, Xiaohan Wang, Yu Wu, Wei Li 等CVPR 2022 · 被引用 58 次
- Towards Real-World Prohibited Item Detection: A Large-Scale X-ray BenchmarkBoying Wang, Libo Zhang, Longyin Wen, Xianglong Liu 等ICCV 2021 · 被引用 110 次
- PANDA: A Gigapixel-Level Human-Centric Video DatasetXueyang Wang, Xiya Zhang, Yinheng Zhu, Yuchen Guo 等CVPR 2020
- Triple-Cooperative Video Shadow DetectionZhihao Chen, Liang Wan, Lei Zhu, Jia Shen 等CVPR 2021
- Characterizing and Detecting Non-Consensual Photo Sharing on Social NetworksTengfei Zheng, Tongqing Zhou, Qiang Liu, Kui Wu 等CCS 2022 · 被引用 6 次
