VidTune: Creating Video Soundtracks with Generative Music and Video-Based Thumbnails
Mina Huh, C. Ailie Fraser, Dingzeyu Li, Mira Dontcheva, Bryan Wang
摘要
Music shapes the tone of videos, yet creators find it hard to find soundtracks that match their video’s mood and narrative. Recent text-to-music models let creators generate music from text prompts, but our formative study (N=8) shows creators struggle to construct diverse prompts, quickly review and compare tracks, and understand their impact on the video. We present VidTune, a system that supports soundtrack creation by generating diverse music options from a creator’s prompt and producing contextual thumbnails for rapid review. VidTune extracts representative video subjects to ground thumbnails in context, maps each track’s valence and energy onto visual cues like color and brightness, and depicts prominent genres and instruments. Creators can refine tracks with natural language edits, which VidTune expands into new generations. In a controlled user study (N=12) and an exploratory case study (N=6), participants found VidTune helpful for efficiently reviewing and comparing music options and described the process as playful and enriching.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Novice-AI Music Co-Creation via AI-Steering Tools for Deep Generative ModelsRyan Louie, Andy Coenen, Cheng Zhi Huang, Michael Terry 等CHI 2020 · 被引用 265 次
- Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language ModelsStephen Brade, Bryan Wang, Maurício Sousa, Sageev Oore 等UIST 2023 · 被引用 179 次
- Luminate: Structured Generation and Exploration of Design Space with Large Language Models for Human-AI Co-CreationSangho Suh, Meng Chen, Bryan Min, Toby Jia-Jun Li 等CHI 2024 · 被引用 143 次
- FashionQ: An AI-Driven Creativity Support Tool for Facilitating Ideation in Fashion DesignYoungseung Jeon, Seungwan Jin, Patrick C. Shih, Kyungsik HanCHI 2021 · 被引用 143 次
相关 Paper
- VidMuse: A Simple Video-to-Music Generation Framework with Long-Short-Term ModelingZeyue Tian, Zhaoyang Liu, Ruibin Yuan, Jiahao Pan 等CVPR 2025
- "Is Text-Based Music Search Enough to Satisfy Your Needs?" A New Way to Discover Music with ImagesJeongeun Park, Hyorim Shin, Changhoon Oh, Ha Young KimCHI 2024 · 被引用 6 次
- MVPrompt: Building Music-Visual Prompts for AI Artists to Craft Music Video Mise-en-scèneChungHa Lee, Daeho Lee, Jin-Hyuk HongCHI 2025 · 被引用 5 次
- Vidmento: Creating Video Stories through Context-Aware Expansion with Generative VideoCatherine Yeh, Anh Truong, Mira Dontcheva, Bryan WangCHI 2026 · 被引用 1 次
- SoundStager: Interactive Design of Story-Driven GenAI Soundscapes for VideoSuhyeon Yoo, Adolfo Hernandez Santisteban, Prem Seetharaman, Justin Salamon 等CHI 2026 · 被引用 2 次
