DIPA2: An Image Dataset with Cross-cultural Privacy Perception Annotations
Anran Xu, Zhongyi Zhou, Kakeru Miyazaki, Ryo Yoshikawa, Simo Hosio, Koji Yatani
Abstract
The world today is increasingly visual. Many of the most popular online social networking services are largely powered by images, making image privacy protection a critical research topic in the fields of ubiquitous computing, usable security, and human-computer interaction (HCI). One topical issue is understanding privacy-threatening content in images that are shared online. This dataset article introduces DIPA2, an open-sourced image dataset that offers object-level annotations with high-level reasoning properties to show perceptions of privacy among different cultures. DIPA2 provides 5,897 annotations describing perceived privacy risks of 3,347 objects in 1,304 images. The annotations contain the type of the object and four additional privacy metrics: 1) information type indicating what kind of information may leak if the image containing the object is shared, 2) a 7-point Likert item estimating the perceived severity of privacy leakages, and 3) intended recipient scopes when annotators assume they are either image owners or allowing others to repost the image. Our dataset contains unique data from two cultures: We recruited annotators from both Japan and the U.K. to demonstrate the impact of culture on object-level privacy perceptions. In this paper, we first illustrate how we designed and performed the construction of DIPA2, along with data analysis of the collected annotations. Second, we provide two machine-learning baselines to demonstrate how DIPA2 challenges the current image privacy recognition task. DIPA2 facilitates various types of research on image privacy, including machine learning methods inferring privacy threats in complex scenarios, quantitative analysis of cultural influences on privacy preferences, understanding of image sharing behaviors, and promotion of cyber hygiene for general user populations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Examining Human Perception of Generative Content Replacement in Image Privacy ProtectionAnran Xu, Shitao Fang, Huan Yang, Simo Hosio et al.CHI 2024 · 22 citations
- What If Smart Homes Could See Our Homes?: Exploring DIY Smart Home Building Experiences with VLM-Based Camera SensorsSojeong Yun, Youn-kyung LimCHI 2025 · 5 citations
- Imago Obscura: An Image Privacy AI Co-pilot to Enable Identification and Mitigation of RisksKyzyl Monteiro, Yuchen Wu, Sauvik DasUIST 2025 · 2 citations
- Collab: Fostering Critical Identification of Deepfake Videos on Social Media via Synergistic AnnotationShuning Zhang, Linzhi Wang, Shixuan Li, Yuanyuan Wu et al.CHI 2026 · 1 citation
- When Privacy Meets Recovery: The Overlooked Half of Surrogate-Driven Privacy Preservation for MLLM EditingSiyuan Xu, Yibing Liu, Peilin Chen, Yung-Hui Li et al.AAAI 2026
Builds on10
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
- Privacy-Enhancing Technology and Everyday Augmented Reality: Understanding Bystanders' Varying Needs for Awareness and ConsentJoseph O'Hagan, Pejman Saeghe, Jan Gugenheimer, Daniel Medeiros et al.UbiComp 2023 · 102 citations
- Towards A Taxonomy of Content Sensitivity and Sharing Preferences for PhotosYifang Li, Nishant Vishwamitra, Hongxin Hu, Kelly CaineCHI 2020 · 48 citations
- Influencing Photo Sharing Decisions on Social Media: A Case of Paradoxical FindingsMary Jean Amon, Rakibul Hasan, Kurt Hugenberg, Bennett I. Bertenthal et al.S&P 2020 · 47 citations
- Your Photo is so Funny that I don't Mind Violating Your Privacy by Sharing it: Effects of Individual Humor Styles on Online Photo-sharing BehaviorsRakibul Hasan, Bennett I. Bertenthal, Kurt Hugenberg, Apu KapadiaCHI 2021 · 31 citations
Related papers
- Human Attributes Prediction under Privacy-preserving ConditionsAnshu Singh, Shaojing Fan, Mohan S. KankanhalliACM MM 2021 · 10 citations
- Private Attribute Inference from Images with Vision-Language ModelsBatuhan Tömekçe, Mark Vero, Robin Staab, Martin T. VechevNeurIPS 2024 · 54 citations
- Everyone's Privacy Matters! An Analysis of Privacy Leakage from Real-World Facial Images on Twitter and Associated User BehaviorsYuqi Niu, Weidong Qiu, Peng Tang, Lifan Wang et al.CSCW 2025 · 3 citations
- User Preferences for Interdependent Privacy Preservation Strategies in Social MediaAaron Necaise, Tangila Islam Tanni, Aneka Williams, Yan Solihin et al.CSCW 2023 · 9 citations
- Behind the Meme: Understanding User Experiences with Memes on Social MediaYuqi Niu, Dilara Keküllüoglu, Weidong Qiu, Nadin KokciyanCHI 2026 · 1 citation
