Uncovering Human Traits in Determining Real and Spoofed Audio: Insights from Blind and Sighted Individuals
Chaeeun Han, Prasenjit Mitra, Syed Masum Billah
摘要
This paper explores how blind and sighted individuals perceive real and spoofed audio, highlighting differences and similarities between the groups. Through two studies, we find that both groups focus on specific human traits in audio–such as accents, vocal inflections, breathing patterns, and emotions–to assess audio authenticity. We further reveal that humans, irrespective of visual ability, can still outperform current state-of-the-art machine learning models in discerning audio authenticity; however, the task proves psychologically demanding. Moreover, detection accuracy scores between blind and sighted individuals are comparable, but each group exhibits unique strengths: the sighted group excels at detecting deepfake-generated audio, while the blind group excels at detecting text-to-speech (TTS) generated audio. These findings not only deepen our understanding of machine-manipulated and neural-renderer audio but also have implications for developing countermeasures, such as perceptible watermarks and human-AI collaboration strategies for spoofing detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Beyond Visual Perception: Insights from Smartphone Interaction of Visually Impaired Users with Large Multimodal ModelsJingyi Xie, Rui Yu, He Zhang, Syed Masum Billah 等CHI 2025 · 被引用 40 次
- Characterizing Photorealism and Artifacts in Diffusion Model-Generated ImagesNegar Kamali, Karyn Nakamura, Aakriti Kumar, Angelos Chatzimparmpas 等CHI 2025 · 被引用 23 次
- SpeakEasy: Enhancing Text-to-Speech Interactions for Expressive Content CreationStephen Brade, Sam Anderson, Rithesh Kumar, Zeyu Jin 等CHI 2025 · 被引用 7 次
- Effect of AI Performance, Risk Perception, and Trust on Human Dependence in Deepfake Detection AI SystemYingfan Zhou, Ester Chen, Manasa Pisipati, Aiping Xiong 等CSCW 2025 · 被引用 2 次
- Characterizing the Impact of Audio Deepfakes in the Presence of Cochlear ImplantMagdalena Pasternak, Kevin Warren, Daniel Olszewski, Susan Nittrouer 等NDSS 2025
它引用的顶会 Paper2
相关 Paper
- Blind and Low-Vision Individuals' Detection of Audio DeepfakesFilipo Sharevski, Aziz Zeidieh, Jennifer Vander Loop, Peter JachimCCS 2024 · 被引用 5 次
- "Better Be Computer or I'm Dumb": A Large-Scale Evaluation of Humans as Audio Deepfake DetectorsKevin Warren, Tyler Tucker, Anna Crowder, Daniel Olszewski 等CCS 2024 · 被引用 9 次
- VoiceRadar: Voice Deepfake Detection using Micro-Frequency and Compositional AnalysisKavita Kumari, Maryam Abbasihafshejani, Alessandro Pegoraro, Phillip Rieger 等NDSS 2025
- Joint Audio-Visual Deepfake DetectionYipin Zhou, Ser-Nam LimICCV 2021 · 被引用 232 次
- Emotions Don't Lie: An Audio-Visual Deepfake Detection Method using Affective CuesTrisha Mittal, Uttaran Bhattacharya, Rohan Chandra, Aniket Bera 等ACM MM 2020 · 被引用 314 次
