Lune

ICML2026顶会

Beyond Pixels: Mining Compressed Domain Artifacts for Efficient AI-Generated Video Detection

Anran Zhu, Zhengli Shi, Chende Zheng, Chenhao Lin, Zhengyu Zhao, Le Yang, Chong Zhang, Shuai Liu, Chao Shen

出版方
2026年份

摘要

With the rapid advancement of high-fidelity video generation models, robust AI-generated video (AIGV) detection has become increasingly needed. While most AIGV detection methods operate in the decoded pixel domain, we observe that detection in the pixel domain inevitably entangles task-irrelevant semantic information, leading to substantial semantic redundancy and extensive redundant computation, while overlooking free-to-use signals in compressed bitstreams. In particular, motion vectors and residuals directly encode temporal and spatial generative artifacts but remain largely underexplored. To address these issues, we propose a unified framework for S patio- T emporal RE sidual and A rtifact M ining, namely STREAM , which enables AIGV detection directly from compressed bitstreams. STREAM leverages I-frames, motion vectors, and residual errors to capture spatiotemporal artifacts that are typically smoothed out by decompression filters. In particular, we design a lightweight network with a motion-guided alignment module and a gated fusion mechanism, enabling adaptive fusion of spatial artifacts and nonlinear temporal dynamics. Extensive experimental results demonstrate that STREAM achieves SOTA performance with an mAP of 0.965, with 2.5× faster inference than previous SOTA baselines.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper10

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖