Record Once, Post Everywhere: Automatic Shortening of Audio Stories for Social Media
Bryan Wang, Zeyu Jin, Gautham J. Mysore
摘要
Following the prevalence of short-form video, short-form voice content has emerged on social media platforms like Twitter and Facebook. A challenge that creators face is hard constraints on the content length. If the initial recording is not short enough, they need to re-record or edit their content. Both are time-consuming, and the latter, if supported, can have a learning curve. Moreover, creators need to manually create multiple versions to publish content on platforms with different length constraints. To simplify this process, we present ROPE1 (Record Once, Post Everywhere). Creators can record voice content once, and our system will automatically shorten it to all length limits by removing parts of the recording for each target. We formulate this as a combinatorial optimization problem and propose a novel algorithm that automatically selects optimal sentence combinations from the original content to comply with each length constraint. Creators can customize the algorithmically shortened content by specifying sentences to include or exclude. Our system can also use the user-specified constraints to recompute and provides a new version. We conducted a user study comparing ROPE with a sentence-based manual editing baseline. The results show that ROPE can generate high-quality edits, alleviating the cognitive loads of creators for shortening content. While our system and user study address short-form voice content specifically, we believe that the same concept can also be applied to other media such as video with narration and dialog.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- VideoDiff: Human-AI Video Co-Creation with AlternativesMina Huh, Ding Li, Kim Pimmel, Hijung Valentina Shin 等CHI 2025 · 被引用 26 次
- Soundify: Matching Sound Effects to VideoDavid Chuan-En Lin, Anastasis Germanidis, Cristóbal Valenzuela, Yining Shi 等UIST 2023 · 被引用 15 次
- GazeNoter: Co-Piloted AR Note-Taking via Gaze Selection of LLM Suggestions to Match Users' IntentionsHsin-Ruey Tsai, Shih-Kang Chiu, Bryan WangCHI 2025 · 被引用 8 次
- SoundStager: Interactive Design of Story-Driven GenAI Soundscapes for VideoSuhyeon Yoo, Adolfo Hernandez Santisteban, Prem Seetharaman, Justin Salamon 等CHI 2026 · 被引用 2 次
- TalkLess: Blending Extractive and Abstractive Summarization for Editing Speech to Preserve Content and StyleKarim Benharrak, Puyuan Peng, Amy PavelUIST 2025 · 被引用 1 次
它引用的顶会 Paper4
- Rescribe: Authoring and Automatically Editing Audio DescriptionsAmy Pavel, Gabriel Reyes, Jeffrey P. BighamUIST 2020 · 被引用 72 次
- Hierarchical Summarization for Longform Spoken DialogDaniel Li, Thomas Chen, Albert Tung, Lydia B. ChiltonUIST 2021 · 被引用 19 次
- StreamHover: Livestream Transcript Summarization and AnnotationSangwoo Cho, Franck Dernoncourt, Tim Ganter, Trung Bui 等EMNLP 2021 · 被引用 18 次
- MemSum: Extractive Summarization of Long Documents Using Multi-Step Episodic Markov Decision ProcessesNianlong Gu, Elliott Ash, Richard H. R. HahnloserACL 2022
相关 Paper
- ReDirector: Creating Any-Length Video Retakes with Rotary Camera EncodingByeongjun Park, Byung-Hoon Kim, Hyungjin Chung, Jong ChulCVPR 2026 · 被引用 10 次
- Making Short-Form Videos Accessible with Hierarchical Video SummariesTess Van Daele, Akhil Iyer, Yuning Zhang, Jalyn C. Derry 等CHI 2024 · 被引用 37 次
- Short Video Ordering via Position Decoding and Successor PredictionShiping Ge, Qiang Chen, Zhiwei Jiang, Yafeng Yin 等SIGIR 2024 · 被引用 1 次
- SpeakEasy: Enhancing Text-to-Speech Interactions for Expressive Content CreationStephen Brade, Sam Anderson, Rithesh Kumar, Zeyu Jin 等CHI 2025 · 被引用 7 次
- A Character-Level Length-Control Algorithm for Non-Autoregressive Sentence SummarizationPuyuan Liu, Xiang Zhang, Lili MouNeurIPS 2022 · 被引用 21 次
