VideoDiff: Human-AI Video Co-Creation with Alternatives
Mina Huh, Ding Li, Kim Pimmel, Hijung Valentina Shin, Amy Pavel, Mira Dontcheva
摘要
To make an engaging video, people sequence interesting moments and add visuals such as B-rolls or text. While video editing requires time and effort, AI has recently shown strong potential to make editing easier through suggestions and automation. A key strength of generative models is their ability to quickly generate multiple variations, but when provided with many alternatives, creators struggle to compare them to find the best fit. We propose VideoDiff, an AI video editing tool designed for editing with alternatives. With VideoDiff, creators can generate and review multiple AI recommendations for each editing process: creating a rough cut, inserting B-rolls, and adding text effects. VideoDiff simplifies comparisons by aligning videos and highlighting differences through timelines, transcripts, and video previews. Creators have the flexibility to regenerate and refine AI suggestions as they compare alternatives. Our study participants (N=12) could easily compare and customize alternatives, creating more satisfying results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment DesignWen-Fan Wang, Ting-Ying Lee, Chien-Ting Lu, Che-Wei Hsu 等UIST 2025 · 被引用 4 次
- VidTune: Creating Video Soundtracks with Generative Music and Video-Based ThumbnailsMina Huh, C. Ailie Fraser, Dingzeyu Li, Mira Dontcheva 等CHI 2026 · 被引用 1 次
- Designing Multi-Robot Ground Video Sensemaking with Public Safety ProfessionalsPuqi Zhou, Ali Asgarov, Aafiya Hussain, Wonjoon Park 等CHI 2026 · 被引用 1 次
- Vidmento: Creating Video Stories through Context-Aware Expansion with Generative VideoCatherine Yeh, Anh Truong, Mira Dontcheva, Bryan WangCHI 2026 · 被引用 1 次
它引用的顶会 Paper33
- AI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model PromptsTongshuang Wu, Michael Terry, Carrie Jun CaiCHI 2022 · 被引用 465 次
- Creating Augmented and Virtual Reality Applications: Current Practices, Challenges, and OpportunitiesNarges Ashtari, Andrea Bunt, Joanna McGrenere, Michael Nebeling 等CHI 2020 · 被引用 274 次
- Co-Writing Screenplays and Theatre Scripts with Language Models: Evaluation by Industry ProfessionalsPiotr Mirowski, Kory W. Mathewson, Jaylen Pittman, Richard EvansCHI 2023 · 被引用 235 次
- The Effects of Generative AI on Design Fixation and Divergent ThinkingSamangi Wadinambiarachchi, Ryan M. Kelly, Saumya Pareek, Qiushi Zhou 等CHI 2024 · 被引用 186 次
- Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language ModelsStephen Brade, Bryan Wang, Maurício Sousa, Sageev Oore 等UIST 2023 · 被引用 179 次
相关 Paper
- SoundStager: Interactive Design of Story-Driven GenAI Soundscapes for VideoSuhyeon Yoo, Adolfo Hernandez Santisteban, Prem Seetharaman, Justin Salamon 等CHI 2026 · 被引用 2 次
- TokenFlow: Consistent Diffusion Features for Consistent Video EditingMichal Geyer, Omer Bar-Tal, Shai Bagon, Tali DekelICLR 2024 · 被引用 439 次
- Structure and Content-Guided Video Synthesis with Diffusion ModelsPatrick Esser, Johnathan Chiu, Parmida Atighehchian, Jonathan Granskog 等ICCV 2023 · 被引用 733 次
- EditBoard: Towards a Comprehensive Evaluation Benchmark for Text-Based Video Editing ModelsYupeng Chen, Penglin Chen, Xiaoyu Zhang, Yixian Huang 等AAAI 2025 · 被引用 5 次
- FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video SynthesisFeng Liang, Bichen Wu, Jialiang Wang, Licheng Yu 等CVPR 2024
