Lune

AAAI2025顶会

EditBoard: Towards a Comprehensive Evaluation Benchmark for Text-Based Video Editing Models

Yupeng Chen, Penglin Chen, Xiaoyu Zhang, Yixian Huang, Qian Xie

2025年份
5被引次数
1顶会引用

摘要

The rapid development of diffusion models has significantly advanced AI-generated content (AIGC), particularly in Textto-Image (T2I) and Text-to-Video (T2V) generation. Textbased video editing, leveraging these generative capabilities, has emerged as a promising field, enabling precise modifications to video content based on textual prompts. Despite the proliferation of innovative video editing models, there is a conspicuous lack of comprehensive evaluation frameworks that holistically assess these models' performance across various dimensions. Existing metrics are limited, inconsistent, and focused on assigning a single score per metric, failing to reveal model's performance on each editing task. To address this gap, we propose EditBoard, the first comprehensive evaluation benchmark for text-based video editing models. EditBoard encompasses nine automatic metrics across four key dimensions, evaluating models on four categories of tasks, and introduces three new metrics to assess fidelity. This task-oriented framework facilitates objective evaluation by breaking down model performance into details, providing insights into each model's strengths and weaknesses. By open-sourcing EditBoard, we aim to standardize evaluation and advance the development of robust video editing models.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper1

问问它们各自怎么用它

它引用的顶会 Paper19

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖