Combating Falsification of Speech Videos with Live Optical Signatures
Hadleigh Schwartz, Xiaofeng Yan, Charles J. Carver, Xia Zhou
Abstract
High-profile speech videos are prime targets for falsification, owing to their accessibility and influence. This work proposes VeriLight, a low-overhead and unobtrusive system for protecting speech videos from visual manipulations of speaker identity and lip and facial motion. Unlike the predominant purely digital falsification detection methods, VeriLight creates dynamic physical signatures at the event site and embeds them into all video recordings via imperceptible modulated light. These physical signatures encode semantically-meaningful features unique to the speech event, including the speaker's identity and facial motion, and are cryptographically-secured to prevent spoofing. The signatures can be extracted from any video downstream and validated against the portrayed speech content to check its integrity. Key elements of VeriLight include (1) a framework for generating extremely compact (i.e., 150-bit), pose-invariant speech video features, based on locality-sensitive hashing; and (2) an optical modulation scheme that embeds 200 bps into video while remaining imperceptible both in video and live. Experiments on extensive video datasets show VeriLight achieves AUCs 0.99 and a true positive rate of 100% in detecting falsified videos. Further, VeriLight is highly robust across recording conditions, video post-processing techniques, and white-box adversarial attacks on its feature extraction methods. A demonstration of VeriLight is available at https://mobilex.cs.columbia.edu/verilight.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on29
- Towards Evaluating the Robustness of Neural NetworksNicholas Carlini, David A. WagnerS&P 2017 · 9,786 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Leveraging Frequency Analysis for Deep Fake Image RecognitionJoel Frank, Thorsten Eisenhofer, Lea Schönherr, Asja Fischer et al.ICML 2020 · 848 citations
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 710 citations
- End-to-End Reconstruction-Classification Learning for Face Forgery DetectionJunyi Cao, Chao Ma, Taiping Yao, Shen Chen et al.CVPR 2022 · 327 citations
Related papers
- Learning Forgery-Aware Lip Representations Without Forgery PriorsBofan Chen, Hongyu Zhu, Yi He, Sichu Liang et al.CVPR 2026
- Identity-Aware Vision-Language Model for Explainable Face Forgery DetectionJunhao Xu, Jingjing Chen, Yang Jiao, Jiacheng Zhang et al.AAAI 2026 · 1 citation
- Lips Don't Lie: A Generalisable and Robust Approach To Face Forgery DetectionAlexandros Haliassos, Konstantinos Vougioukas, Stavros Petridis, Maja PanticCVPR 2021
- EchoFence: Non-Intrusive Forgery Detection in Video Conferencing via Ultrasonic SensingLeqi Zhao, Luxin Shi, Jianwei Liu, Rui Xiao et al.INFOCOM 2026
- Ariadne's Thread of LipSync: Unraveling Forgeries via Inconsistency between Lip Motions and Head PosesTianyi She, Jiawei Liu, Weifeng Liu, Hanqing Zhao et al.ICML 2026
