AI-Generated Video Detection via Perceptual Straightening
Christian Internò, Robert Geirhos, Markus Olhofer, Sunny Liu, Barbara Hammer, David A. Klindt
Abstract
The rapid advancement of generative AI enables highly realistic synthetic videos, posing significant challenges for content authentication and raising urgent concerns about misuse. Existing detection methods often struggle with generalization and capturing subtle temporal inconsistencies. We propose ReStraV(Representation Straightening for Video), a novel approach to distinguish natural from AI-generated videos. Inspired by the "perceptual straightening" hypothesis [1, 2]-which suggests real-world video trajectories become more straight in neural representation domain-we analyze deviations from this expected geometric property. Using a pre-trained self-supervised vision transformer (DINOv2), we quantify the temporal curvature and stepwise distance in the model's representation domain. We aggregate statistics of these measures for each video and train a classifier. Our analysis shows that AI-generated videos exhibit significantly different curvature and distance patterns compared to real videos. A lightweight classifier achieves state-of-the-art detection performance (e.g., 97.17% accuracy and 98.63% AUROC on the VidProM benchmark [3]), substantially outperforming existing image-and video-based methods. ReStraV is computationally efficient, offering a low-cost and effective detection solution. This work provides new insights into using neural representation geometry for AI-generated video detection. Classifier (e.g., MLP) In representation SSE space, natural videos trace straighter paths than AIgenerated videos. The trajectory geometry provides a discriminative signal. SSE (e.g., DINOv2) Frames are processed by a SSE and we collect the embeddings. Classifier: AI-generated vs. natural Trajectories in representation domain AI-Generated AI-Generated vs. Natural Natural .. .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2e887907-ff95-44e9-a751-bf28d5781c1cCited by top-tier papers5
- Skyra: AI-Generated Video Detection via Grounded Artifact ReasoningYifei Li, Wenzhao Zheng, Yanran Zhang, Runze Sun et al.CVPR 2026 · 24 citations
- Temporal Straightening for Latent PlanningYing Wang, Oumayma Bounou, Gaoyue Zhou, Randall Balestriero et al.ICML 2026 · 19 citations
- VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement LearningHao Tan, jun lan, Senyuan Shi, Zichang Tan et al.ICML 2026 · 12 citations
- Training-free Detection of Generated Videos via Spatial-Temporal LikelihoodsOmer Ben Hayun, Roy Betser, Meir Yossef Levi, Levi Kassel et al.CVPR 2026 · 7 citations
- Explainable Forensics of Manipulated Segments in Untrimmed Long VideosYue Feng, Jingjing Li, Qijia Lu, Wei Ji et al.ICML 2026
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Physics-Driven Spatiotemporal Modeling for AI-Generated Video DetectionShuhai Zhang, Zihao Lian, Jiahao Yang, Daiyuan Li et al.NeurIPS 2025 · 29 citations
- D3: Training-Free AI-Generated Video Detection Using Second-Order FeaturesChende Zheng, Ruiqi Suo, Chenhao Lin, Zhengyu Zhao et al.ICCV 2025 · 11 citations
- Detecting Generated Images by Fitting Natural Image DistributionsYonggang Zhang, Jun Nie, Xinmei Tian, Mingming Gong et al.NeurIPS 2025 · 9 citations
- Preserving Forgery Artifacts: AI-Generated Video Detection at Native ScaleZhengcen Li, Chenyang Jiang, Hang Zhao, Shiyang Zhou et al.ICLR 2026 · 8 citations
- Learning predictable and robust neural representations by straightening image sequencesXueyan Niu, Cristina Savin, Eero P. SimoncelliNeurIPS 2024 · 13 citations
