SpeechMirror: A Multimodal Visual Analytics System for Personalized Reflection of Online Public Speaking Effectiveness
Ze-Yuan Huang, Qiang He, Kevin T. Maher, Xiaoming Deng, Yu-Kun Lai, Cuixia Ma, Sheng Feng Qin, Yong-Jin Liu, Hongan Wang
Abstract
As communications are increasingly taking place virtually, the ability to present well online is becoming an indispensable skill. Online speakers are facing unique challenges in engaging with remote audiences. However, there has been a lack of evidence-based analytical systems for people to comprehensively evaluate online speeches and further discover possibilities for improvement. This paper introduces SpeechMirror, a visual analytics system facilitating reflection on a speech based on insights from a collection of online speeches. The system estimates the impact of different speech techniques on effectiveness and applies them to a speech to give users awareness of the performance of speech techniques. A similarity recommendation approach based on speech factors or script content supports guided exploration to expand knowledge of presentation evidence and accelerate the discovery of speech delivery possibilities. SpeechMirror provides intuitive visualizations and interactions for users to understand speech factors. Among them, SpeechTwin, a novel multimodal visual summary of speech, supports rapid understanding of critical speech factors and comparison of different speech samples, and SpeechPlayer augments the speech video by integrating visualization of the speaker's body language with interaction, for focused analysis. The system utilizes visualizations suited to the distinct nature of different speech factors for user comprehension. The proposed system and visualization techniques were evaluated with domain experts and amateurs, demonstrating usability for users with low visualization literacy and its efficacy in assisting users to develop insights for potential improvement.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7205a3f2-ae89-4dd5-98da-7c076c6130b0Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Spatially Conditioned Graphs for Detecting Human-Object InteractionsFrederic Z. Zhang, Dylan Campbell, Stephen GouldICCV 2021 · 170 citations
- Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise TransformerFrederic Z. Zhang, Dylan Campbell, Stephen GouldCVPR 2022 · 118 citations
- VoiceCoach: Interactive Evidence-based Training for Voice Modulation Skills in Public SpeakingXingbo Wang, Haipeng Zeng, Yong Wang, Aoyu Wu et al.CHI 2020 · 34 citations
- E-ffective: A Visual Analytic System for Exploring the Emotion and Effectiveness of Inspirational SpeechesKevin T. Maher, Ze-Yuan Huang, Jian-Cheng Song, Xiaoming Deng et al.IEEE VIS 2021 · 15 citations
Related papers
- RealityTalk: Real-Time Speech-Driven Augmented Presentation for AR Live StorytellingJian Liao, Adnan Karim, Shivesh Singh Jadon, Rubaiat Habib Kazi et al.UIST 2022 · 44 citations
- VisAug: Facilitating Speech-Rich Web Video Navigation and Engagement with Auto-Generated Visual AugmentationsBaoquan Zhao, Xiaofan Ma, Qianshi Pang, Ruomei Wang et al.ACM MM 2025 · 1 citation
- AffectiveSpotlight: Facilitating the Communication of Affective Responses from Audience Members during Online PresentationsPrasanth Murali, Javier Hernandez, Daniel McDuff, Kael Rowan et al.CHI 2021 · 81 citations
- "I Am a Mirror Dweller": Probing the Unique Strategies Users Take to Communicate in the Context of Mirrors in Social Virtual RealityKexue Fu, Yixin Chen, Jiaxun Cao, Xin Tong et al.CHI 2023 · 48 citations
- Glass Chirolytics: Reciprocal Compositing and Shared Gestural Control for Face-to-Face Collaborative Visualization at a DistanceDion Barja, Matthew BrehmerCHI 2026 · 1 citation
