Video Analysis and Generation via a Semantic Progress Function
Gal Metzer, Sagi Polaczek, Ali Mahdavi-Amiri, Raja Giryes, Daniel Cohen-Or
摘要
Transformations produced by image and video generation models often evolve in a highly non-linear manner: long stretches where the content barely changes are followed by sudden, abrupt semantic jumps. To analyze and correct this behavior, we introduce a Semantic Progress Function, a one-dimensional representation that captures how the meaning of a given sequence evolves over time. For each frame, we compute distances between semantic embeddings and fit a smooth curve that reflects the cumulative semantic shift across the sequence. Departures of this curve from a straight line reveal uneven semantic pacing. Building on this insight, we propose a semantic linearization procedure that reparameterizes (or retimes) the sequence so that semantic change unfolds at a constant rate, yielding smoother and more coherent transitions. Beyond linearization, our framework provides a model-agnostic foundation for identifying temporal irregularities, comparing semantic pacing across different generators, and steering both generated and real-world video sequences toward arbitrary target pacing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Sigmoid Loss for Language Image Pre-TrainingXiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, Lucas BeyerICCV 2023 · 被引用 2,932 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- ReStyle: A Residual-Based StyleGAN Encoder via Iterative RefinementYuval Alaluf, Or Patashnik, Daniel Cohen-OrICCV 2021 · 被引用 377 次
- Temporally Consistent Transformers for Video GenerationWilson Yan, Danijar Hafner, Stephen James, Pieter AbbeelICML 2023 · 被引用 47 次
相关 Paper
- Versatile Transition Generation with Image-to-Video DiffusionZuhao Yang, Jiahui Zhang, Yingchen Yu, Shijian Lu 等ICCV 2025
- Unsupervised Coherent Video Cartoonization with Perceptual Motion ConsistencyZhenhuan Liu, Liang Li, Huajie Jiang, Xin Jin 等AAAI 2022 · 被引用 7 次
- Anchoring and Rescaling Attention for Semantically Coherent InbetweeningTae Eun Choi, Sumin Shim, Junhyeok Kim, Seong Jae HwangCVPR 2026 · 被引用 2 次
- From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute TransitionLing Lo, Kelvin C. K. Chan, Wen-Huang Cheng, Ming-Hsuan YangICCV 2025 · 被引用 3 次
- Pathways on the Image Manifold: Image Editing via Video GenerationNoam Rotstein, Gal Yona, Daniel Silver, Roy Velich 等CVPR 2025
