Synthesis-Assisted Video Prototyping From a Document
Peggy Chi, Tao Dong, Christian Früh, Brian Colonna, Vivek Kwatra, Irfan Essa
摘要
Video productions commonly start with a script, especially for talking head videos that feature a speaker narrating to the camera. When the source materials come from a written document – such as a web tutorial, it takes iterations to refine content from a text article to a spoken dialogue, while considering visual compositions in each scene. We propose Doc2Video, a video prototyping approach that converts a document to interactive scripting with a preview of synthetic talking head videos. Our pipeline decomposes a source document into a series of scenes, each automatically creating a synthesized video of a virtual instructor. Designed for a specific domain – programming cookbooks, we apply visual elements from the source document, such as a keyword, a code snippet or a screenshot, in suitable layouts. Users edit narration sentences, break or combine sections, and modify visuals to prototype a video in our Editing UI. We evaluated our pipeline with public programming cookbooks. Feedback from professional creators shows that our method provided a reasonable starting point to engage them in interactive scripting for a narrated instructional video.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- DataParticles: Block-based and Language-oriented Authoring of Animated Unit VisualizationsYining Cao, Jane L. E, Chen Zhu-Tian, Haijun XiaCHI 2023 · 被引用 53 次
- VideoDiff: Human-AI Video Co-Creation with AlternativesMina Huh, Ding Li, Kim Pimmel, Hijung Valentina Shin 等CHI 2025 · 被引用 26 次
- Papeos: Augmenting Research Papers with Talk VideosTae Soo Kim, Matt Latzke, Jonathan Bragg, Amy X. Zhang 等UIST 2023 · 被引用 17 次
- Compositional Structures as Substrates for Human-AI Co-creation Environment: A Design Approach and A Case StudyYining Cao, Yiyi Huang, Anh Truong, Hijung Valentina Shin 等CHI 2025 · 被引用 15 次
- Reflecting on Design Paradigms of Animated Data Video ToolsLeixian Shen, Haotian Li, Yun Wang, Huamin QuCHI 2025 · 被引用 10 次
它引用的顶会 Paper14
- Rescribe: Authoring and Automatically Editing Audio DescriptionsAmy Pavel, Gabriel Reyes, Jeffrey P. BighamUIST 2020 · 被引用 72 次
- Automatic Generation of Two-Level Hierarchical Tutorials from Instructional Makeup VideosAnh Truong, Peggy Chi, David Salesin, Irfan Essa 等CHI 2021 · 被引用 57 次
- Towards Supporting Programming Education at Scale via Live StreamingYan Chen, Walter S. Lasecki, Tao DongCSCW 2020 · 被引用 45 次
- Crosscast: Adding Visuals to Audio Travel PodcastsHaijun Xia, Jennifer Jacobs, Maneesh AgrawalaUIST 2020 · 被引用 44 次
- RubySlippers: Supporting Content-based Voice Navigation for How-to VideosMinsuk Chang, Mina Huh, Juho KimCHI 2021 · 被引用 40 次
相关 Paper
- Automatic Instructional Video Creation from a Markdown-Formatted TutorialPeggy Chi, Nathan Frey, Katrina Panovich, Irfan EssaUIST 2021 · 被引用 27 次
- Code2Video: A Code-centric Paradigm for Educational Video CreationYanzhe Chen, Kevin Qinghong Lin, Mike Zheng ShouICML 2026
- Automatic Video Creation From a Web PagePeggy Chi, Zheng Sun, Katrina Panovich, Irfan EssaUIST 2020 · 被引用 25 次
- PaperTok: Exploring the Use of Generative AI for Creating Short-form Videos for Research CommunicationMeziah Ruby Cristobal, Hyeonjeong Byeon, Tze-Yu Chen, Ruoxi Shang 等CHI 2026 · 被引用 2 次
- Demo2Tutorial: From Human Experience to Multimodal Software TutorialsZechen Bai, Zhiheng Chen, Yiqi Lin, Kevin Qinghong Lin 等CVPR 2026
