Improving Automatic Summarization for Browsing Longform Spoken Dialog
Daniel Li, Thomas Chen, Alec Zadikian, Albert Tung, Lydia B. Chilton
摘要
Longform spoken dialog delivers rich streams of informative content through podcasts, interviews, debates, and meetings. While production of this medium has grown tremendously, spoken dialog remains challenging to consume as listening is slower than reading and difficult to skim or navigate relative to text. Recent systems leveraging automatic speech recognition (ASR) and automatic summarization allow users to better browse speech data and forage for information of interest. However, these systems intake disfluent speech which causes automatic summarization to yield readability, adequacy, and accuracy problems. To improve navigability and browsability of speech, we present three training agnostic post-processing techniques that address dialog concerns of readability, coherence, and adequacy. We integrate these improvements with user interfaces which communicate estimated summary metrics to aid user browsing heuristics. Quantitative evaluation metrics show a 19% improvement in summary quality. We discuss how summarization technologies can help people browse longform audio in trustworthy and readable ways.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- Rambler: Supporting Writing With Speech via LLM-Assisted Gist ManipulationSusan Lin, Jeremy Warner, J. D. Zamfirescu-Pereira, Matthew G. Lee 等CHI 2024 · 被引用 38 次
- Making Short-Form Videos Accessible with Hierarchical Video SummariesTess Van Daele, Akhil Iyer, Yuning Zhang, Jalyn C. Derry 等CHI 2024 · 被引用 37 次
- EmBARDiment: an Embodied AI Agent for Productivity in XRRiccardo Bovo, Steven Abreu, Karan Ahuja, Eric J. Gonzalez 等IEEE VR 2025 · 被引用 19 次
- GazeNoter: Co-Piloted AR Note-Taking via Gaze Selection of LLM Suggestions to Match Users' IntentionsHsin-Ruey Tsai, Shih-Kang Chiu, Bryan WangCHI 2025 · 被引用 8 次
- A Sound Understanding - An In-Situ Deployment of an Accessible Audio-Media Player with People Living with AphasiaFilip Bircanin, Alexandre Nevsky, Madeline N. Cruice, Ognjen Markovic 等CHI 2026 · 被引用 1 次
相关 Paper
- Hierarchical Summarization for Longform Spoken DialogDaniel Li, Thomas Chen, Albert Tung, Lydia B. ChiltonUIST 2021 · 被引用 19 次
- Towards Abstractive Grounded Summarization of Podcast TranscriptsKaiqiang Song, Chen Li, Xiaoyang Wang, Dong Yu 等ACL 2022 · 被引用 11 次
- Desirable Unfamiliarity: Insights from Eye Movements on Engagement and Readability of Dictation InterfacesZhaohui Liang, Yonglin Chen, Naser Al Madi, Can LiuCHI 2026 · 被引用 1 次
- Methods for Evaluating the Fluency of Automatically Simplified Texts with Deaf and Hard-of-Hearing Adults at Various Literacy LevelsOliver Alonzo, Jessica Trussell, Matthew Watkins, Sooyeon Lee 等CHI 2022 · 被引用 8 次
- SummN: A Multi-Stage Summarization Framework for Long Input Dialogues and DocumentsYusen Zhang, Ansong Ni, Ziming Mao, Chen Henry Wu 等ACL 2022
